Anthropic released Claude Opus 4.8 on May 28, positioning honesty and reliability — not benchmark gains — as the flagship improvements, alongside a new fast mode and dynamic workflows that let Claude Code coordinate hundreds of parallel subagents.
The honesty pitch
The most quoted line from the launch: Opus 4.8 is roughly four times less likely than its predecessor to let flaws in code it has written pass unremarked. For teams deploying AI agents on production codebases, silent failure is the failure mode that matters — an agent that flags its own uncertainty is worth more than one that scores higher while hiding its mistakes.
Fast mode and the pricing menu
Opus 4.8 keeps its predecessor's pricing at $5 per million input tokens and $25 per million output, with a 1 million token context window. The new fast mode — up to 2.5x faster — costs double. Anthropic is effectively letting customers price their own latency tolerance, a small design choice that says a lot about how mature the API business has become.
Dynamic workflows
The headline capability for engineering organizations: a single Claude Code session can now plan and run hundreds of subagents in parallel, enabling codebase-scale migrations across hundreds of thousands of lines. Available on the Claude API, Amazon Bedrock, Google Vertex AI, Microsoft Foundry, GitHub Copilot, and GitLab from day one — distribution breadth that has quietly become one of Anthropic's structural advantages.
