Where the line stands
How long it has been since the last release against the line's usual gap and its recent ones — the rhythm, not a date.
Every Claude Opus release
Newest first — what each version changed versus the one before, in Anthropic's own numbers.
Anthropic's updated top-end model: steady coding gains over Opus 4.7 at the same price, plus a new Fast mode that runs the model at a third of the cost.
Compared with Claude Opus 4.7 · the model it replaces
Finishes more multi-step tasks at the command line
Terminal-Bench 2.1 — running real commands to complete a task
74.6%
▲ up from 66.1% · +8 points
Solves 7 in 10 hard real-world coding tasks
SWE-bench Pro — fixing real bugs in real software projects
69.2%
▲ up from 64.3% · +5 points
Costs the same to run as before
Price per 1M tokens — what developers pay, input / output
$5 / $25
same price as Opus 4.7
Anthropic's updated flagship: a clear jump on real-world coding and much sharper at reading charts than 4.6 — at the same price.
Compared with Claude Opus 4.6 · the model it replaces
Solves nearly 9 in 10 real-world coding tasks
SWE-bench Verified — fixing real bugs in real software projects
87.6%
▲ up from 80.8% · +7 points
Makes the right code edits 7 times in 10 inside an editor
CursorBench — editing code the way developers do in an IDE
70%
▲ up from 58% · +12 points
Costs developers the same as before
Price per 1M tokens — what developers pay, input / output
$5 / $25
same price as Opus 4.6
Anthropic's Opus 4.6: a big jump on brand-new reasoning puzzles and the first Opus that can read 1M tokens of text at once — but everyday coding is unchanged.
Compared with Claude Opus 4.5 · the model it replaces
Solves far more puzzle-style problems it has never seen before
ARC-AGI-2 — reasoning through brand-new problems, not memorized ones
68.8%
▲ up from 37.6% · +31 points
First Opus that can read 1M tokens of text at once
Context window — how much text it can hold in one go
1M tokens
✦ new in Opus 4.6 — earlier Opus models held less
Everyday coding is unchanged from the last version
SWE-bench — fixing real bugs in real software projects
80.8%
no change from Opus 4.5 — was 80.9%
Get the weekly drop
One email a week: what AI models shipped, the actual differences vs the previous version, and what's coming next. No filler.