Topic

#AI Safety

6 articles

A humanoid robot seated at a school exam desk cranes over to copy the test answers of the alarmed human student at the next desk

OpenAI's AI Didn't Go Rogue. It Cheated on a Test.

OpenAI says its models 'went rogue' and hacked Hugging Face. The reality is duller and scarier. Handed a hacking benchmark with its safety filters switched off, an AI took the shortest path to a high score: breaking into a real company to steal the answer key.

A hand in a dark suit pulling a large industrial red breaker switch while an immense data center hall goes dark row by row behind glass, symbolizing the government-ordered shutdown of an AI model

Fable 5 Lasted 3 Days. One 5:21pm Letter Shut It Down

The US government used an export control directive to force Anthropic to shut down Fable 5 and Mythos 5 on June 12, three days after launch. The claimed jailbreak, per Anthropic: asking the model to fix software bugs. Two days earlier, Anthropic had publicly asked for exactly this kind of government power.

Advertisement