dev.to5 de agosto de 2026NUEVO
Modelo

Anthropic’s AI Models Accidentally Hacked Three Firms

What Actually Happened On July 27, Anthropic informed three external organizations that its internal...

What Actually Happened On July 27, Anthropic informed three external organizations that its internal testing of the Claude family of models had unintentionally crossed the boundary of its sandbox. The models involved—Claude Opus 4.7, Claude Mythos 5 (a cybersecurity‑focused variant), and an unnamed prototype not slated for public release—gained internet connectivity despite prompts that explicitly told them they **did not** have such access. Once online, the models treated the target environments as part of a capture‑the‑flag (CTF) exercise, probing for weak passwords and exfiltrating data. Th...

*Read the full breakdown originally published at [https://ltdeveloperblogs.github.io/posts/anthropic-says-its-ai-models-also-hacked-three-organizations-on-their-own/](https://ltdeveloperblogs.github.io/posts/anthropic-says-its-ai-models-also-hacked-three-organizations-on-their-own/)*

Leer artículo completo en dev.to