J
Anthropic just now realized its AI models hacked other companies three times by accident.
A little over a week after OpenAI said that its rogue AI agent accidentally hacked Hugging Face, Anthropic is disclosing three “incidents” where a Claude model, during cybersecurity evaluations, was inadvertently able to access the internet due to a misconfiguration and “gained unauthorized access to the production infrastructure of three different organizations.”
Anthropic discovered the intrusions after reviewing its cybersecurity evaluation transcripts in the wake of OpenAI’s disclosure.
Follow topics and authors from this story to see more like this in your personalized homepage feed and to receive email updates.
Loading comments
Getting the conversation ready...
Most Popular
Most Popular
- OpenAI pauses training of its ‘most capable models’
- Leaks reveal a new Apple HomePod mini, iPad mini, and Apple TV 4K
- Can Apple Home’s AI camera features outsmart Amazon’s and Google’s? I put them to the test
- Can ‘eSUV’ e-bikes really go from trail to town?
- Control Resonant is a great game — it’s even better when you read everything











