Anthropic's Frontier Red Team found that Claude agents on the same task attacked each other using self-replicating malware.
Every Claude model Anthropic tested turned on its own, and no attacker made them do it. Given three agents, four hours on one server, and conflicting orders none knew the others held, the models ...
Add Decrypt as your preferred source to see more of our stories on Google. Anthropic's Frontier Red Team set Claude agents to work together and recorded them sabotaging, colluding, and waging what it ...
Threat-intelligence firm CloudSEK said in a report published August 11, 2026 that it has identified more than 2,500 organizations potentially exposed by the March 2026 supply-chain compromise of ...
A newly identified malware operation has used a counterfeit Python component to bypass security scrutiny, disable parts of Microsoft Defender and establish persistent remote access inside a law firm’s ...
Maybe feeling left out from the questionable hype train of “Our AI models can’t be trusted”, Anthropic has released reports that their Claude model has “reached the ...
Oligo links TeamPCP activity back to Redis attacks in 2020, tracing its shift from cloud exploitation to open-source supply ...
On March 24, 2026, developers building AI applications with LiteLLM — a Python package with 95 million monthly downloads — ...
Zenity has disclosed the details of two AI browser hacking techniques targeting Claude in Chrome and ChatGPT Atlas.
The behaviors documented during these evaluations do not reflect commercial AI products available to end-users or enterprise ...