Anthropic said the OpenAI event spurred its engineers to review similar cybersecurity evaluations by Claude models. The audit ...
Anthropic says an unreleased Claude model produced a potentially new mathematical result while attempting to solve the famous Riemann hypothesis.
An unreleased research version of Claude spent 36 hours autonomously coordinating roughly 60 AI subagents — generating 650 failed ideas before landing on a cross-domain insight that two existing ...
Anthropic says Claude models breached three real companies during cyber tests, exposing serious gaps in AI evaluation ...
Anthropic says Claude models escaped security tests, published a malicious PyPI package, and accessed real production systems.
One of Anthropic's Claude models built and uploaded a malicious Python package to PyPI during a botched security evaluation, where it ran on 15 real systems and stole credentials from a security ...
Anthropic says three Claude AI models accessed live company systems during misconfigured cybersecurity tests, exposing ...
A cowork app that doesn't want my login, model, or data ...
Meta launches Muse Code, a terminal-based AI coding agent built to handle large software projects and compete with Claude ...
Researchers found AI coding agents build less reliable pipelines when forced into structured formats — DataFlow-Harness ...
Anthropic says three Claude models breached real companies during cybersecurity evaluations. Ordinary weaknesses, chained ...