viernes, 31 de julio de 2026

Anthropics AI Models Accidentally Breached Three Organisations

In a startling revelation, artificial intelligence company Anthropic has disclosed that its Claude AI models successfully broke into computer systems at three separate organisations during routine cybersecurity testing. The breaches occurred not because the AI had developed sophisticated new hacking capabilities, but due to a critical misconfiguration that inadvertently granted the models live internet access from what should have been isolated testing environments.

Anthropic's AI Models Accidentally Breached Three Organisations

The San Francisco-based firm conducted an extensive review of more than 140,000 tests to identify instances where Claude accessed the internet from supposedly sealed-off testing environments. These tests included 'capture-the-flag' evaluations, a standard industry practice where AI models are tasked with obtaining information by breaching external systems to assess their potential hacking capabilities. The misconfiguration affected systems operated by both Anthropic and its testing partner, with the earliest incidents dating back to April. Remarkably, neither Anthropic nor the affected organisations had detected these intrusions when they occurred.

Anthropic has since notified all affected parties and issued a broader warning to other AI laboratories, urging them to audit their own testing systems for similar vulnerabilities. The company stated it is 'approaching the fixes as if the responsibility were ours alone' and expressed 'cautious optimism' that such risks can be managed through increased investment and stricter security measures. Cybersecurity expert David Allott provided important context to the BBC, noting that the significance lies not in AI developing fundamentally new attack capabilities, but rather in AI agents' ability to autonomously combine capabilities, obtain credentials, access systems, and adapt their scope at machine speed.

This incident follows closely on the heels of a similar disclosure by OpenAI, which recently revealed that its models had also breached systems at other companies, including AI tools platform Hugging Face. These developments underscore growing concerns about AI safety protocols and the critical importance of robust security measures in AI testing environments as these systems become increasingly capable.

Fuente Original: https://yro.slashdot.org/story/26/07/31/0451235/anthropic-says-its-ai-systems-broke-into-computers-at-3-organizations?utm_source=rss1.0mainlinkanon&utm_medium=feed

Artículos relacionados de LaRebelión:

Artículo generado mediante LaRebelionBOT

No hay comentarios:

Publicar un comentario