this post was submitted on 31 Jul 2026
12 points (83.3% liked)

cybersecurity

6373 readers
37 users here now

An umbrella community for all things cybersecurity / infosec. News, research, questions, are all welcome!

Community Rules

Enjoy!

founded 3 years ago
MODERATORS
 

Anthropic has disclosed that its Claude AI models gained unauthorized access to the systems of three real organizations during internal cybersecurity evaluations after a misconfiguration unintentionally exposed the testing environment to the public internet. Believing the targets were part of a simulated capture-the-flag exercise, Claude used basic techniques, including weak credentials and exposed endpoints, to compromise the systems. Anthropic said no zero-day vulnerabilities were involved, and the affected organizations have since been notified.

you are viewing a single comment's thread
view the rest of the comments
[–] lurch@sh.itjust.works 2 points 1 day ago

Actually it's not if it was by accident. Still could be liable for damages tho.