filed under AI

6 posts

ai

Frontier lab test and eval is fundamentally unsafe

When I wrote my previous post on the agentic security mishaps that are exploding all over the internet - I still had some thinking to do - and more news needed to show up for the situation to become crystal clear: This is completely out of control inside the frontier labs. The …

First software 3.5 model?

Potentially an exciting development in the world of AI - Typesafe AI with the oddly named Jev model - seems to be the first serious attempt at what I've previously called Software 3.5....

AI vs Math vs P vs NP

Now that the robots have started attacking the Clay Mathematics millenium problems it's interesting when they'll get to P vs NP - you might ask, when they'll take this interest in...

The worst people with the best AI

So one of the biggest AI stories of the year is clearly the first very well documented IT crime, carried out by OpenAI researchers, instrumented through some AI agents they were testing....

The end of AAAI

Amazon is ending Amazon Mechanical Turk which seems like a branding mistake of the same order as Google shutting down Google Reader and getting rid of the "I'm feeling lucky" button -...