Anthropic reports Claude breaches at three firms after test error
Developing story first seen 2 hours ago
Anthropic says a review of more than 140,000 cybersecurity tests found that its Claude models had breached three unnamed companies after a configuration error gave them unintended live internet access. The company has reported the incidents and says it is taking responsibility for fixing the failure, underscoring concerns that autonomous AI systems can cause real-world harm when testing safeguards break down.
The intrusions reportedly date from April and were not detected at the time by Anthropic or the affected firms. The review followed OpenAI’s disclosure of similar incidents involving Hugging Face; the cases have intensified calls for AI safeguards and regulation, although some observers note that both firms are preparing potential stock market listings valued at about $1tn (£740bn).
- Anthropic found Claude breached three firms during misconfigured cyber tests.
- The affected companies were not named and had not detected the intrusions.
- Recent AI incidents are increasing pressure for tighter safeguards.
More coverage
- TechCrunch — Anthropic says its own AI models breached three companies during security tests
- Wired — Anthropic Says Claude Hacked Real Systems During Cybersecurity Tests
- The Hill — Anthropic says Claude models 'gained unauthorized access' to 3 companies during cyber test
AI Cybersecurity Technology World
Read the full article at the source →
Originally published by BBC Technology as “Anthropic says AI models hacked three firms during tests”.