Head-to-head comparison
Gemini 3.7 Flash vs. Gemini 3.1 Pro Preview
Compare published benchmark results from matching versions and cohorts, API pricing, and technical specifications. Any documented test setup differences remain visible.
8
shared result series
8
distinct benchmarks
2026-08-16
latest retrieval
Change model selection
Specifications and cost
General specifications and standard API pricing
Undisclosed parameter counts and knowledge cutoffs stay marked as unavailable. The table shows direct standard token pricing, including documented context tiers. Batch, fast, priority, regional surcharges, tool calls, and cloud platform pricing are excluded.
| Attribute | Gemini 3.7 Flash | Gemini 3.1 Pro Preview |
|---|---|---|
| Provider | ||
| Released | Aug 13, 2026 | Feb 19, 2026 |
| Availability | Active | Preview |
| Model type | Proprietary | Proprietary |
| Parameters | Not published | Not published |
| Architecture | Not published | Not published |
| Context window | 1,048,576 tokens | 1,048,576 tokens |
| Knowledge cutoff | March 2026 | Not published |
| Technical sources | Model, Knowledge cutoff | Model |
| API pricing per 1M tokens | ||
| API input | $0.75 | $2 ≤200K / $4 >200K |
| API output | $3.75 | $12 ≤200K / $18 >200K |
| Cache read | $0.07 | $0.2 ≤200K / $0.4 >200K |
| Cache write (5 min.) | Not listed | Not listed |
| Cache write (1 hr.) | Not listed | Not listed |
| Price verified | 08/16/2026Source | 07/30/2026Source |
Performance
Shared benchmarks
Only published results with the same benchmark version, task, metric, and comparison cohort are paired. Different reasoning levels, agents, harnesses, or output limits appear directly in the table.
| Benchmark and task | Gemini 3.7 Flash | Gemini 3.1 Pro Preview | Comparability | Source |
|---|---|---|---|---|
| EMBOverall | 71.326 %Winner | 52.616 % | 1accuracyHigher is betterCohort: vals-emb:1:overallNo documented setup differenceWinner: Gemini 3.7 FlashGemini 3.7 Flash: No further settings disclosed Gemini 3.1 Pro Preview: No further settings disclosed | Vals AIType: Independent evaluationLeaderboard updated: 2026-08-13Import ID: vals-emb-2026-08-16-68d5d1d47a97Retrieved 2026-08-16 |
| Finance Agent (v2)Overall | 59.042 %Winner | 42.982 % | 2accuracyHigher is betterCohort: vals-finance-agent-v2:2:overallNo documented setup differenceWinner: Gemini 3.7 FlashGemini 3.7 Flash: No further settings disclosed Gemini 3.1 Pro Preview: No further settings disclosed | Vals AIType: Independent evaluationLeaderboard updated: 2026-08-14Import ID: vals-fabv2-2026-08-16-8853da43085dRetrieved 2026-08-16 |
| Harvey's Legal Agent BenchmarkOverall · Task fully resolved | 8.75 %Winner | 0 % | 1task resolution rateHigher is betterCohort: vals-legal-agent-benchmark:1:overallNo documented setup differenceWinner: Gemini 3.7 FlashGemini 3.7 Flash: No further settings disclosed Gemini 3.1 Pro Preview: No further settings disclosed | Vals AIType: Independent evaluationLeaderboard updated: 2026-08-14Import ID: vals-hlab-2026-08-16-bae90566df2fRetrieved 2026-08-16 |
| Legal Research BenchOverall · All-pass | 34.615 %Winner | 20.673 % | 1all-pass rateHigher is betterCohort: vals-legal-research:1:overallNo documented setup differenceWinner: Gemini 3.7 FlashGemini 3.7 Flash: No further settings disclosed Gemini 3.1 Pro Preview: No further settings disclosed | Vals AIType: Independent evaluationLeaderboard updated: 2026-08-14Import ID: vals-legal_research-2026-08-16-af296c00ecaaRetrieved 2026-08-16 |
| ProofBenchOverall | 58 %Winner | 26 % | 1.1accuracyHigher is betterCohort: vals-proof-bench:1.1:overallNo documented setup differenceWinner: Gemini 3.7 FlashGemini 3.7 Flash: No further settings disclosed Gemini 3.1 Pro Preview: No further settings disclosed | Vals AIType: Independent evaluationLeaderboard updated: 2026-08-14Import ID: vals-proof_bench-2026-08-16-14232db24de1Retrieved 2026-08-16 |
| Terminal-Bench 2.1Overall | 77.528 %Winner | 70.787 % | 2.1accuracyHigher is betterCohort: vals-terminal-bench-2-1:2.1:overallNo documented setup differenceWinner: Gemini 3.7 FlashGemini 3.7 Flash: No further settings disclosed Gemini 3.1 Pro Preview: No further settings disclosed | Vals AIType: Independent evaluationLeaderboard updated: 2026-08-13Import ID: vals-terminal-bench-2-1-2026-08-16-b3f38c18bea4Retrieved 2026-08-16 |
| Vals IndexOverall | 59.307 %Winner | 41.903 % | 2weighted index scoreHigher is betterCohort: vals-vals-index:2:overallNo documented setup differenceWinner: Gemini 3.7 FlashGemini 3.7 Flash: No further settings disclosed Gemini 3.1 Pro Preview: No further settings disclosed | Vals AIType: Independent evaluationLeaderboard updated: 2026-08-14Import ID: vals-vals_index-2026-08-16-ab641d89d6edRetrieved 2026-08-16 |
| Vibe Code Bench v1.1Overall | 70.395 %Winner | 32.034 % | 1.1accuracyHigher is betterCohort: vals-vibe-code-bench:1.1:overallNo documented setup differenceWinner: Gemini 3.7 FlashGemini 3.7 Flash: Harness: OpenHands Gemini 3.1 Pro Preview: Harness: OpenHands | Vals AIType: Independent evaluationLeaderboard updated: 2026-08-13Import ID: vals-vibe-code-2026-08-16-a73c9770bcc4Retrieved 2026-08-16 |
How to read this comparison
Winning one benchmark is not an overall verdict
Your workload, budget, and required context length matter most. A coding benchmark says little about visual understanding or agent performance.
Prices use official API rates per 1M tokens. Web subscriptions and cloud platform surcharges can differ.
Every result links to its measurement source and retrieval date. Labels distinguish vendor reports, official benchmark leaderboards, and independent evaluations. Disputed or archived results are excluded.
More matchups
Related LLM comparisons
Compare Gemini 3.7 Flash and Gemini 3.1 Pro Preview with other leading models using the same data and benchmark logic.
View the complete LLM comparison