Claude Opus 5
Anthropic
- Released
- July 24, 2026
- Data date
- July 30, 2026
Category view
Position within the category
This overview uses only published data from matching cohorts. Missing values never change a rank.
The index averages rank percentiles from 4 documented comparison cohorts. Only models with complete coverage receive a position.
- Position
- Rank 1 of 26
- Index score
- 88.9 / 100
- Coverage
- 4 / 4
Leaderboard
Measurements
Comparable benchmark results
Each chart contains exactly one source, one measurement series, and one stored comparison cohort. Bars show the position. The measured value appears on the right.
SimpleQA Verified v2
63.3% · Rank 4 of 5
Task: Kaggle score · Comparison cohort: simpleqa-verified-v2:official:overall · Data date: August 9, 2026
5 of 5 model versions shown in this chart. A higher value ranks first.
1,000 verified prompts without tools. Kaggle independently reproduced the results.
AlmanBench v0.1
94.95% · Rank 3 of 5
Task: Accepted cases · Comparison cohort: almanbench-v0-1:official:overall · Data date: August 9, 2026
5 of 5 model versions shown in this chart. A higher value ranks first.
1,029 public tasks with one direct run per row. The target is the Alman language specification.
Individual values
Published individual values
No exactly matching published comparison cohort is available for these values. The bar shows only the documented scale, not a rank.
ARC-AGI-3
Task: RHAE overall score · Data date: August 26, 2026
Scale 0 to 100. Not ranked
Official ARC Prize harness on unseen interactive environments. Reasoning levels remain separate.
Profile
Specifications and access
Published information about this model. Unknown values are not estimated.
- Model type
- ProprietarySource
- Context window
- 1,000,000 tokensSource
- Knowledge cutoff
- May 2026Source
- Notes
- Released July 24, 2026. Near-frontier intelligence at launch, close to Fable 5 at half the price ($5/$25). 1M context, 128K output, and a May 2026 knowledge cutoff. It became the new default on Claude Max and strongest model on Claude Pro at launch. It led Frontier-Bench v0.1, surpassed Fable 5 on OSWorld 2.0, came within 0.5% of Fable 5 on CursorBench 3.2, and trailed Mythos 5 on cybersecurity. Those results describe the pre-Fable-5.1 and pre-Mythos-5.1 cohort.Source
Pricing
Published prices
Prices remain tied to their documented unit and source.
Head-to-head comparisons
Compare this model
Each matchup compares this model with exactly one other model from the same category.
More models
Models from the same selection
All AI models
Evidence
Primary sources and data date
Every statement links to its underlying documentation or leaderboard.