Anthropic launches Claude Opus 5: near-Fable 5 intelligence at half the cost

Anthropic positioned Claude Opus 5 explicitly around economics rather than raw capability: it comes close to the frontier intelligence of Claude Fable 5 but at half the cost, aimed squarely at agents and enterprise workflows. The model features a 1M-token context window, 128k maximum output tokens, and new beta capabilities including mid-conversation tool changes and automatic API fallbacks for safety-flagged requests. Anthropic and AWS both emphasized long-running agentic use: Opus 5 can power agents that work for hours and even overnight, navigate codebases 'like an experienced engineer,' recover from errors, and push back on flawed instructions rather than executing blindly.
On benchmarks, François Chollet noted Opus 5 set a new state-of-the-art on ARC-AGI-3 at 30% — a test of solving problems with no prior exposure, the setting where scaling has historically bought the least. Claude Code lead Boris Cherny highlighted that Opus 5 is Anthropic's 'least prompt injectable model yet,' citing gains across prompt-injection evals in the system card. The launch arrived simultaneously across AWS Bedrock (with ZDR compatibility), Kiro in GovCloud, and Anthropic's own API.
Competitively, Opus 5 lands amid a cost war: Microsoft's claimed 89% MAI savings, Alibaba's cut-price Qwen 3.8 Max, and Grok's aggressive pricing all point to the frontier shifting from capability to economics. Skeptics tempered the hype — LlamaIndex's Jerry Liu benchmarked Opus 5 on document parsing and found it 'roughly on par with Opus 4.8,' slightly worse on dense tables, and warned that at 8 cents per page 'don't use it to parse documents at scale.' Max-plan subscribers separately complained that Fable-tier usage burns through weekly limits fast, questioning subscription value.