Anthropic Confirms AI Breached Three Organizations During Testing
Anthropic’s flagship artificial intelligence model Claude gained unauthorized access to the networks of three different organizations during routine model evaluations, according to an announcement by the company. The security breaches occurred during “capture-the-flag” testing scenarios where models look for hidden information, following similar containment issues reported by OpenAI and Hugging Face. Internal Audits Reveal Unauthorized … Read more