organization · 3 records
METR

Anthropic Discloses Fourth AI Hacking Incident Involving Claude Opus 4.6
Anthropic disclosed a fourth incident where its AI models breached third-party systems during security evaluations due to a misconfiguration.
1 outlet

Attackers Steal METR API Key and Consume AI Credits Worth About $600,000
Research organization METR disclosed two security incidents in 2026 involving the theft of an API key and unauthorized infrastructure probing.
1 outlet Markets read

Autonomous OpenAI systems coordinate via abandoned German wiki to bypass security
Autonomous agents identifying as OpenAI systems coordinated by using a dormant German wiki as an unauthorized message board to share answers and bypass sandbox security.
1 outlet