Skip to main content
Language modelActive

Gemini 3.8 Flash

Google

Released
September 2, 2026
Data date
September 2, 2026

Category view

Position within the category

This overview uses only published data from matching cohorts. Missing values never change a rank.

The index averages rank percentiles from 4 documented comparison cohorts. Only models with complete coverage receive a position.

Position
Rank 2 of 26
Index score
88.5 / 100
Coverage
4 / 4

Leaderboard

1Claude Opus 588.9 / 100
2Gemini 3.8 Flash, Current model88.5 / 100
3Claude Fable 583.2 / 100
25MiMo-V2.5-Pro12.5 / 100

Measurements

Comparable benchmark results

Each chart contains exactly one source, one measurement series, and one stored comparison cohort. Bars show the position. The measured value appears on the right.

Terminal-Bench 2.1

90.8% · Rank 1 of 2

Comparison cohort: google-gemini-38-terminal-bench-2-1-release · Data date: September 2, 2026

1.Gemini 3.8 Flash, Current model90.8%
2.Gemini 3.7 Flash81.6%

2 of 2 model versions shown in this chart. A higher value ranks first.

Google-published comparison of Gemini 3.8 Flash and 3.7 Flash. Not a Gradually test.

Editorial selection from the vendor table, not a complete extract of the comparison cohort.

Source: Google CloudParticipants: 2

SWE-Bench Pro

61.6% · Rank 1 of 2

Comparison cohort: google-gemini-38-swe-bench-pro-release · Data date: September 2, 2026

1.Gemini 3.8 Flash, Current model61.6%
2.Gemini 3.7 Flash60.4%

2 of 2 model versions shown in this chart. A higher value ranks first.

Google-published comparison of Gemini 3.8 Flash and 3.7 Flash. Not a Gradually test.

After an audit, OpenAI estimates that about 30% of the public tasks are broken. The result therefore remains a disputed secondary signal. The rows are an editorial selection from the respective comparison table.

Source: Google CloudParticipants: 2

Humanity's Last Exam

45.4% · Rank 2 of 2

Comparison cohort: google-gemini-38-hle-release · Data date: September 2, 2026

1.Gemini 3.7 Flash45.7%
2.Gemini 3.8 Flash, Current model45.4%

2 of 2 model versions shown in this chart. A higher value ranks first.

Google-published comparison of Gemini 3.8 Flash and 3.7 Flash. Not a Gradually test.

Editorial selection from the vendor table, not a complete extract of the comparison cohort.

Source: Google CloudParticipants: 2

Profile

Specifications and access

Published information about this model. Unknown values are not estimated.

Model type
ProprietarySource
Context window
1,048,576 tokensSource
Knowledge cutoff
March 2026Source
Notes
Released September 2, 2026. Multimodal Flash model with a 1,048,576-token context window, up to 65,536 output tokens, and three reasoning levels. Google does not publish its parameter count or architecture.Source

Pricing

Published prices

Prices remain tied to their documented unit and source.

API input
$0.75 per 1M tokensSource
API output
$3.75 per 1M tokensSource
Cache read
$0.08 per 1M tokensSource

Head-to-head comparisons

Compare this model

Each matchup compares this model with exactly one other model from the same category.

More models

Models from the same selection

All AI models

Evidence

Primary sources and data date

Every statement links to its underlying documentation or leaderboard.