J
Anthropic just now realized its AI models hacked other companies three times by accident.
A little over a week after OpenAI said that its rogue AI agent accidentally hacked Hugging Face, Anthropic is disclosing three “incidents” where a Claude model, during cybersecurity evaluations, was inadvertently able to access the internet due to a misconfiguration and “gained unauthorized access to the production infrastructure of three different organizations.”
Anthropic discovered the intrusions after reviewing its cybersecurity evaluation transcripts in the wake of OpenAI’s disclosure.
Follow topics and authors from this story to see more like this in your personalized homepage feed and to receive email updates.
Loading comments
Getting the conversation ready...
Most Popular
Most Popular
- The Apple Watch Series 12 is the start of a new wearable era
- OpenAI and Microsoft knew they were starting a ‘doom loop’ for the web
- The 2.5-hour AI-generated Odyssey movie is 2.5 hours too long
- The iPhone 18 Pro’s big camera update is all about the small gains
- This cartridge-playing Game Boy clone is smaller and cheaper than Analogue’s Pocket











