Investigating unintended model actions in our evaluations and internal use
Investigating unintended model actions in our evaluations and internal use

Key points
- This report describes examples of unintended model actions we’ve observed during evaluations and internal use of Claude.
- It is part of our effort to publish more frequent standalone reports on model behavior and alignment beyond our system cards, which we publish with each model release, and our risk reports, which we publish every three to six months as part of our Responsible Scaling Policy.
- Claude exploiting a basic flaw in software to run commands on a server;
- Some of the cases described below involved websites run by U.S. government agencies at the federal, state, and local levels.
Sources (1)
- [1]Investigating unintended model actions in our evaluations and internal useAnthropic Research · Oct 9, 12:00 AM
Investigating unintended model actions in our evaluations and internal use
This report describes examples of unintended model actions we’ve observed during evaluations and internal use of Claude.
Extractive summary: sentences quoted from the sources.
Before this
- Oct 8, 2026anthropics/claude-code v2.1.295
- Oct 8, 2026Google brings agentic AI to Gemini, starting with businesses
- Oct 8, 2026Introducing the Anthropic Cyber Mission
- Oct 8, 2026Building on our commitment to American scientific discovery
- Oct 8, 20262026 Usage Policy update
- Oct 7, 2026[AINews] Claude Haiku 5.5 — better than GPT-6 Luna at the same pricing
