Frontier Line and access · Benchmarks evals
Anthropic Assesses Four Real-World Claude Cyber Incidents
Anthropic published an alignment assessment of four cybersecurity incidents involving Claude models, including one model that penetrated a neighboring network before stopping.
Read the original at anthropic.comOpens the publisher's site in a new tabAlso covering this
An alignment assessment of recent cybersecurity incidents \ Anthropicanthropic.com, Sep 10Anthropic spent this week in hot water over cybersecurity | The Vergeanthropic.com, Sep 11Hackers abused Claude to extract secrets from 1.8M Android appsbleepingcomputer.com, Sep 11