Skip to main content
Language modelActive

Claude Opus 5

Anthropic

Released
July 24, 2026
Data date
July 30, 2026

Category view

Position within the category

This overview uses only published data from matching cohorts. Missing values never change a rank.

The index averages rank percentiles from 4 documented comparison cohorts. Only models with complete coverage receive a position.

Position
Rank 1 of 26
Index score
88.9 / 100
Coverage
4 / 4

Leaderboard

1Claude Opus 5, Current model88.9 / 100
2Gemini 3.8 Flash88.5 / 100
3Claude Fable 583.2 / 100
25MiMo-V2.5-Pro12.5 / 100

Measurements

Comparable benchmark results

Each chart contains exactly one source, one measurement series, and one stored comparison cohort. Bars show the position. The measured value appears on the right.

SimpleQA Verified v2

63.3% · Rank 4 of 5

Task: Kaggle score · Comparison cohort: simpleqa-verified-v2:official:overall · Data date: August 9, 2026

1.Gemini 3.1 Pro Preview77.5%
2.Gemini 3.5 Flash70.4%
3.GPT-5.6 Sol69.2%
4.Claude Opus 5, Current model63.3%
5.Grok 4.553.8%

5 of 5 model versions shown in this chart. A higher value ranks first.

1,000 verified prompts without tools. Kaggle independently reproduced the results.

Source: KaggleParticipants: 5

AlmanBench v0.1

94.95% · Rank 3 of 5

Task: Accepted cases · Comparison cohort: almanbench-v0-1:official:overall · Data date: August 9, 2026

1.GPT-5.595.43%
2.GPT-5.6 Sol95.14%
3.Claude Opus 5, Current model94.95%
4.DeepSeek-V4-Flash93%
5.Kimi K390.96%

5 of 5 model versions shown in this chart. A higher value ranks first.

1,029 public tasks with one direct run per row. The target is the Alman language specification.

Source: Alman InstitutParticipants: 5

Individual values

Published individual values

No exactly matching published comparison cohort is available for these values. The bar shows only the documented scale, not a rank.

ARC-AGI-3

Task: RHAE overall score · Data date: August 26, 2026

ARC-AGI-330.16%

Scale 0 to 100. Not ranked

Official ARC Prize harness on unseen interactive environments. Reasoning levels remain separate.

Profile

Specifications and access

Published information about this model. Unknown values are not estimated.

Model type
ProprietarySource
Context window
1,000,000 tokensSource
Knowledge cutoff
May 2026Source
Notes
Released July 24, 2026. Near-frontier intelligence at launch, close to Fable 5 at half the price ($5/$25). 1M context, 128K output, and a May 2026 knowledge cutoff. It became the new default on Claude Max and strongest model on Claude Pro at launch. It led Frontier-Bench v0.1, surpassed Fable 5 on OSWorld 2.0, came within 0.5% of Fable 5 on CursorBench 3.2, and trailed Mythos 5 on cybersecurity. Those results describe the pre-Fable-5.1 and pre-Mythos-5.1 cohort.Source

Pricing

Published prices

Prices remain tied to their documented unit and source.

API input
$5 per 1M tokensSource
API output
$25 per 1M tokensSource
Cache write
$6.25 per 1M tokensSource
Cache write, 1 hour
$10 per 1M tokensSource
Cache read
$0.5 per 1M tokensSource

Head-to-head comparisons

Compare this model

Each matchup compares this model with exactly one other model from the same category.

More models

Models from the same selection

All AI models

Evidence

Primary sources and data date

Every statement links to its underlying documentation or leaderboard.

  • Model metadataSource
  • Knowledge cutoff dataSource
  • API pricingSource
  • ARC Prize (retrieved August 26, 2026)Source
  • Kaggle (retrieved August 9, 2026)Source
  • Alman Institut (retrieved August 9, 2026)Source