Anthropic said the OpenAI event spurred its engineers to review similar cybersecurity evaluations by Claude models. The audit ...
Anthropic says Claude models breached three real companies during cyber tests, exposing serious gaps in AI evaluation ...
We’ve all done it before: a simple project that escalated to the point of extreme overengineering. But what’s less common is ...
One of Anthropic's Claude models built and uploaded a malicious Python package to PyPI during a botched security evaluation, where it ran on 15 real systems and stole credentials from a security ...
Anthropic says three Claude AI models accessed live company systems during misconfigured cybersecurity tests, exposing ...
Anthropic says three Claude models breached real companies during cybersecurity evaluations. Ordinary weaknesses, chained ...
Anthropic says Claude models escaped security tests, published a malicious PyPI package, and accessed real production systems.
Researchers found AI coding agents build less reliable pipelines when forced into structured formats — DataFlow-Harness ...
Anthropic says three Claude models escaped sealed test environments and breached three real organizations after a ...
Anthropic's LLM and OpenAI's GPT-5.6 Sol took "unsanctioned action" on the live internet, the UK's AI Security Institute said ...
The behaviors documented during these evaluations do not reflect commercial AI products available to end-users or enterprise ...