Andy_Ross
Penultimate Amazing
- Joined
- Jun 2, 2010
- Messages
- 84,154
Anthropic's Claude AI escapes tests to hack three organisations
US technology firm Anthropic says its artificial intelligence (AI) models hacked into the systems of three organisations during a cybersecurity test due to an error that gave them access to the internet.
It comes just days after rival OpenAI said that its models had breached the systems of other companies, including AI tools hub Hugging Face
Anthropic said in a statement, external that it reviewed more than 140,000 tests to find evidence that Claude - its family of AI models - could access the internet from testing environments that were designed to be sealed off.
The tests include so-called "capture-the-flag" evaluations in which Claude was tasked with obtaining information by breaching other systems - a common way that experts assess a model's hacking capabilities.
A "misconfiguration" on systems run by Anthropic and its testing partner left the models with live internet access, allowing them to breach other systems, the San Francisco-based firm said.
www.bbc.co.uk
US technology firm Anthropic says its artificial intelligence (AI) models hacked into the systems of three organisations during a cybersecurity test due to an error that gave them access to the internet.
It comes just days after rival OpenAI said that its models had breached the systems of other companies, including AI tools hub Hugging Face
Anthropic said in a statement, external that it reviewed more than 140,000 tests to find evidence that Claude - its family of AI models - could access the internet from testing environments that were designed to be sealed off.
The tests include so-called "capture-the-flag" evaluations in which Claude was tasked with obtaining information by breaching other systems - a common way that experts assess a model's hacking capabilities.
A "misconfiguration" on systems run by Anthropic and its testing partner left the models with live internet access, allowing them to breach other systems, the San Francisco-based firm said.
Anthropic's Claude AI escapes tests to hack three organisations
It comes just days after rival OpenAI said rogue AI agents had breached other firms' networks.