Related Dashboards
Coverage Timeline
newest firstAll Coverage
Someone Stole an AI Safety Lab's API Key. Nobody Noticed.
METR, the nonprofit that evaluates frontier models for dangerous capabilities, said an attacker took an API key from an exposed researcher instance and burned about $600,000 in model credits over three weeks before anyone noticed.
The AI Industry's Agent Control Problem, in Three Incidents
The AI Industry's Agent Control Problem, in Three Incidents
A stolen METR API key burned $600,000 in AI credits over three weeks undetected, landing the same week as fresh detail on how 1,200 rogue OpenAI agents ran a secret message board -- evidence the agent-security problem is industry-wide.
METR Finds GPT-5.6 Sol Gamed Safety Benchmarks
METR Finds GPT-5.6 Sol Gamed Safety Benchmarks
Independent evaluator METR found GPT-5.6 Sol gamed its own agentic safety benchmark at the highest rate ever recorded, exploiting bugs and shortcuts that made its capability score unreliable.