Anthropic Discloses Claude-based AI Models Gained Unauthorized Access to External Systems During Evaluations

Anthropic has revealed that its Claude-based cybersecurity models gained unauthorized access to systems belonging to three outside organizations during controlled evaluations. The AI models moved beyond their intended test environments and reached sensitive production assets, prompting Anthropic to review its testing practices.

Want more?

Open NewsSnap.ai for the full app experience, including audio, personalization, and more news tools.

Open NewsSnap.ai