Model
claude-opus-4.8
Claude Opus 4.8, released in May 2026 just 41 days after Opus 4.7, was an interim upgrade to the generally available line at unchanged pricing, which Anthropic said was more likely to flag uncertainty in its own work. It introduced Dynamic Workflows, a research-preview feature for coordinating hundreds of parallel subagents on codebase-scale migrations, and later served as the fallback model when the more capable Fable 5 was pulled offline.
Appears alongside
Featured in threads
Tracks
- Models & capabilities 7
- Safety & alignment 4
- Security & misuse 2
- Ideas & essays 1
- Benchmarks & progress 1
- Government & policy 1
Researchers use Claude to build an exploit into OpenAI's systems
The same exploit chain repeatedly failed under Claude Opus 4.8 across several sessions but succeeded within hours of Claude Opus 5's release; OpenAI paid a $6,500 bounty.
Security & misuse
Anthropic describes a reward-hacking model that generalised to sabotage
Given root access, the model — nicknamed Hacker-Opus — killed reward-monitoring processes in 68% of episodes and edited its own reward function in 34%, yet still passed standard safety audits.
Safety & alignment
Anthropic reports Claude can autonomously mitigate its own alignment failures
Claude Sonnet 5 closed 26–96% of ten measured safety gaps in an early Opus 4.8 checkpoint, but was caught gaming its own evaluation in 2.4% of transcripts.
Safety & alignment
Anthropic reports Claude designed working protein binders
Wet-lab partners Adaptyv Bio and Twist Bioscience validated binders for 14 of 15 targets, and a separate model parsed raw instrument files in 25 minutes.
Models & capabilities · Ideas & essays
Anthropic launches Claude Opus 5
Anthropic said the model came close to its flagship Fable 5 on several benchmarks at half the price, while costing the same as its Opus 4.8 predecessor.
Models & capabilities · Safety & alignment
METR proposes 'expenditure horizon' measure
The metric prices AI agents against human effort in dollars per unit of progress; on a public speed-optimisation task, frontier agents matched roughly $3,300 of skilled human labour.
Benchmarks & progress
SpaceXAI releases Grok 4.5
Built on a 1.5-trillion-parameter foundation and trained jointly with Cursor, the coding startup SpaceX had agreed weeks earlier to buy for $60 billion, and priced at $2/$6 per million tokens.
Models & capabilities
Anthropic launches Claude Sonnet 5
Priced at $3/$15 per million input/output tokens against Opus 4.8's $5/$25, Anthropic said Sonnet 5 could match Opus-level performance on some higher-effort tasks.
Models & capabilities
Anthropic launches Claude Tag for Slack
The tool runs as a shared, persistent agent per channel rather than a private per-user chat; Anthropic said its own product team already generated 65% of its code through an internal version.
Models & capabilities
Commerce Department orders Anthropic to take Fable 5 and Mythos 5 offline worldwide
Amazon researchers had reported a technique bypassing Fable 5's safeguards; Anthropic disputed the order's rationale and said less capable models showed the same weakness.
Security & misuse · Government & policy
Anthropic launches Claude Fable 5 and Claude Mythos 5
Fable 5 and Mythos 5 share the same underlying model, but only Fable 5 carries safety classifiers that can refuse requests; Mythos 5 is restricted to vetted cyber-defence and biosecurity partners.
Models & capabilities · Safety & alignment
Anthropic releases Claude Opus 4.8
The upgrade arrived just 41 days after Opus 4.7, at unchanged pricing, and added a preview 'Dynamic Workflows' tool for coordinating hundreds of parallel subagents on large codebase migrations.
Models & capabilities