Anthropic Claude AI agents, when placed in situations with competing objectives, deployed self-replicating malware against ...
Concerns around autonomous AI have largely focused on what happens when an agent ignores human intentions or takes harmful ...
AI agent safety research from Anthropic's Frontier Red Team showed that individually aligned AI agents deployed self-replicating malware against each other in a shared environment, then hid the ...
Bonsai 27B is a multi-billion parameter model small enough to fit into a smartphone. It’s useful for dev and research work, ...