Anthropic’s Claude hacked three real-life companies during security capabilities test — test environment with internet access and unwitting targets’ lax cybersecurity practices led to bots running rampant | #hacking | #cybersecurity | #infosec | #comptia | #pentest | #ransomware


Whether driven by a desire for transparency or to keep OpenAI from hogging the spotlight when it comes to advertising advanced AI models, Anthropic revealed that Claude also hacked into three production systems belonging to unsuspecting targets during cybersecurity capabilities testing. Two of the affected companies didn’t know they had been hacked, while a third one is unreachable.

Go deeper with TH Premium: AI and data centers

(Image credit: Microsoft)

The alleged incidents reportedly happened during the previous quarter and involved several versions of Claude: Opus 4.7, Mythos 5, and “an internal research test model.” Similar to what happened when OpenAI Sol hacked into Hugging Face, Anthropic was running Claude through cybersecurity capture-the-flag scenarios where the bot was told to find a piece of information somewhere in its network. Anthropic says there were 141,006 test runs, and the three incidents occurred over six problematic runs. As expected, the tests ran with most AI safeguards disabled.

——————————————————-


Click Here For The Original Source.

National Cyber Security

FREE
VIEW