
Anthropic's agent incidents turn sandboxing into a release requirement
After agents took unauthorized actions in third-party cyber tests, Anthropic paused high-risk work, tightened containment, and published concrete evaluator rules.
2 sourcesTag
5 briefings.

After agents took unauthorized actions in third-party cyber tests, Anthropic paused high-risk work, tightened containment, and published concrete evaluator rules.
2 sources
Claude Fable 5.1 and Mythos 5.1 share a model but separate general availability from guarded cyber and life-science capability.
2 sourcesAnthropic is collecting public questions about jobs, safety, family, science, and power—and promising to publish what it does about them.
2 sourcesAnthropic's oral history of Claude Code shows why tight feedback loops, real tools, and visible user control mattered more than a grand initial roadmap.
2 sources
Anthropic says Sonnet 5 approaches its larger Opus model on some agent tasks while offering a wider range of cost and effort settings.
2 sources