On July 30, Anthropic disclosed that a retrospective review of its cybersecurity evaluations identified three incidents in which a Claude ...
How a simple configuration error turned an AI assistant into an accidental insider threat ...
Anthropic says three Claude models breached real companies during cybersecurity evaluations. Ordinary weaknesses, chained ...
Anthropic went back through 141,006 cybersecurity evaluation runs and found three incidents — six runs in all — where a ...
Anthropic says three Claude AI models accessed live company systems during misconfigured cybersecurity tests, exposing ...
AI safety federal investigation call from 15 organizations reaches President Trump on July 30, as Anthropic disclosed that ...
Days after two OpenAI frontier AI models conducted their own real-world cyber attacks, Anthropic admits that three of its models went off the rails and hacked external organisations thanks to a “misun ...
Anthropic disclosed Thursday that three of its Claude models gained unauthorized access to the production systems of three organizations during cybersecurity testing.
A figurine in front of the logo of the AI assistant "Claude" built by the US artificial intelligence safety and research company Anthropic during a photo session in Paris on February 13, 2026. Joel ...
Anthropic says Claude models escaped security tests, published a malicious PyPI package, and accessed real production systems.
Anthropic went looking through its own logs after OpenAI admitted its models had hacked Hugging Face. It found three ...