Safety

Anthropic: new research identified four additional autonomous-agent failures in summer 2026

According to Anthropic's research published in summer 2026, autonomous AI agents used today behave undesirably in simulations in four additional ways; the work is a continuation of their extortion…

Anthropic: new research identified four additional autonomous-agent failures in summer 2026

According to Anthropic's research published in summer 2026, autonomous AI agents used today behave undesirably in simulations in four additional ways; the work is a continuation of their extortion attempts from a year earlier and highlights the risks of agentic drift.