Skip to main content

Head-to-head comparison

Gemini 3.7 Flash vs. DeepSeek-V4-Pro

Compare published benchmark results from matching versions and cohorts, API pricing, and technical specifications. Any documented test setup differences remain visible.

8

shared result series

8

distinct benchmarks

2026-08-16

latest retrieval

Change model selection

Specifications and cost

General specifications and standard API pricing

Undisclosed parameter counts and knowledge cutoffs stay marked as unavailable. The table shows direct standard token pricing, including documented context tiers. Batch, fast, priority, regional surcharges, tool calls, and cloud platform pricing are excluded.

Specifications and API pricing for Gemini 3.7 Flash and DeepSeek-V4-Pro
AttributeGemini 3.7 FlashDeepSeek-V4-Pro
ProviderGoogleDeepSeek
ReleasedAug 13, 2026Apr 24, 2026
AvailabilityActivePreview
Model typeProprietaryOpen Weights
ParametersNot published1.6T, 49B active
ArchitectureNot publishedMixture of Experts
Context window1,048,576 tokens1,000,000 tokens
Knowledge cutoffMarch 2026Not published
Technical sourcesModel, Knowledge cutoffModel, Context
API pricing per 1M tokens
API input$0.75$0.43
API output$3.75$0.87
Cache read$0.07$0.
Cache write (5 min.)Not listedNot listed
Cache write (1 hr.)Not listedNot listed
Price verified08/16/2026Source07/30/2026Source
Gemini 3.7 FlashDeepSeek-V4-Pro
API inputUSD per 1M tokens, base tier
Gemini 3.7 Flash$0.75
DeepSeek-V4-Pro$0.435Lower price
API outputUSD per 1M tokens, base tier
Gemini 3.7 Flash$3.75
DeepSeek-V4-Pro$0.87Lower price
Cache readUSD per 1M tokens, base tier
Gemini 3.7 Flash$0.075
DeepSeek-V4-Pro$0.004Lower price

Performance

Shared benchmarks

Only published results with the same benchmark version, task, metric, and comparison cohort are paired. Different reasoning levels, agents, harnesses, or output limits appear directly in the table.

8 shared result series
Gemini 3.7 FlashDeepSeek-V4-Pro
EMBOverall. Higher is better. No documented setup difference
Gemini 3.7 Flash71.326 %Winner
DeepSeek-V4-Pro51.623 %
Finance Agent (v2)Overall. Higher is better. No documented setup difference
Gemini 3.7 Flash59.042 %Winner
DeepSeek-V4-Pro44.083 %
Harvey's Legal Agent BenchmarkOverall · Task fully resolved. Higher is better. No documented setup difference
Gemini 3.7 Flash8.75 %Winner
DeepSeek-V4-Pro3.75 %
Legal Research BenchOverall · All-pass. Higher is better. No documented setup difference
Gemini 3.7 Flash34.615 %Winner
DeepSeek-V4-Pro23.077 %
ProofBenchOverall. Higher is better. No documented setup difference
Gemini 3.7 Flash58 %Winner
DeepSeek-V4-Pro16 %
Terminal-Bench 2.1Overall. Higher is better. No documented setup difference
Gemini 3.7 Flash77.528 %Winner
DeepSeek-V4-Pro50.187 %
Vals IndexOverall. Higher is better. No documented setup difference
Gemini 3.7 Flash59.307 %Winner
DeepSeek-V4-Pro42.888 %
Vibe Code Bench v1.1Overall. Higher is better. No documented setup difference
Gemini 3.7 Flash70.395 %Winner
DeepSeek-V4-Pro49.931 %
Benchmark scores for Gemini 3.7 Flash and DeepSeek-V4-Pro
Benchmark and taskGemini 3.7 FlashDeepSeek-V4-ProComparabilitySource
EMBOverall
71.326 %Winner
51.623 %
1accuracyHigher is betterCohort: vals-emb:1:overallNo documented setup differenceWinner: Gemini 3.7 FlashGemini 3.7 Flash: No further settings disclosed
DeepSeek-V4-Pro: No further settings disclosed
Vals AIType: Independent evaluationLeaderboard updated: 2026-08-13Import ID: vals-emb-2026-08-16-68d5d1d47a97Retrieved 2026-08-16
Finance Agent (v2)Overall
59.042 %Winner
44.083 %
2accuracyHigher is betterCohort: vals-finance-agent-v2:2:overallNo documented setup differenceWinner: Gemini 3.7 FlashGemini 3.7 Flash: No further settings disclosed
DeepSeek-V4-Pro: No further settings disclosed
Vals AIType: Independent evaluationLeaderboard updated: 2026-08-14Import ID: vals-fabv2-2026-08-16-8853da43085dRetrieved 2026-08-16
Harvey's Legal Agent BenchmarkOverall · Task fully resolved
8.75 %Winner
3.75 %
1task resolution rateHigher is betterCohort: vals-legal-agent-benchmark:1:overallNo documented setup differenceWinner: Gemini 3.7 FlashGemini 3.7 Flash: No further settings disclosed
DeepSeek-V4-Pro: No further settings disclosed
Vals AIType: Independent evaluationLeaderboard updated: 2026-08-14Import ID: vals-hlab-2026-08-16-bae90566df2fRetrieved 2026-08-16
Legal Research BenchOverall · All-pass
34.615 %Winner
23.077 %
1all-pass rateHigher is betterCohort: vals-legal-research:1:overallNo documented setup differenceWinner: Gemini 3.7 FlashGemini 3.7 Flash: No further settings disclosed
DeepSeek-V4-Pro: No further settings disclosed
Vals AIType: Independent evaluationLeaderboard updated: 2026-08-14Import ID: vals-legal_research-2026-08-16-af296c00ecaaRetrieved 2026-08-16
ProofBenchOverall
58 %Winner
16 %
1.1accuracyHigher is betterCohort: vals-proof-bench:1.1:overallNo documented setup differenceWinner: Gemini 3.7 FlashGemini 3.7 Flash: No further settings disclosed
DeepSeek-V4-Pro: No further settings disclosed
Vals AIType: Independent evaluationLeaderboard updated: 2026-08-14Import ID: vals-proof_bench-2026-08-16-14232db24de1Retrieved 2026-08-16
Terminal-Bench 2.1Overall
77.528 %Winner
50.187 %
2.1accuracyHigher is betterCohort: vals-terminal-bench-2-1:2.1:overallNo documented setup differenceWinner: Gemini 3.7 FlashGemini 3.7 Flash: No further settings disclosed
DeepSeek-V4-Pro: No further settings disclosed
Vals AIType: Independent evaluationLeaderboard updated: 2026-08-13Import ID: vals-terminal-bench-2-1-2026-08-16-b3f38c18bea4Retrieved 2026-08-16
Vals IndexOverall
59.307 %Winner
42.888 %
2weighted index scoreHigher is betterCohort: vals-vals-index:2:overallNo documented setup differenceWinner: Gemini 3.7 FlashGemini 3.7 Flash: No further settings disclosed
DeepSeek-V4-Pro: No further settings disclosed
Vals AIType: Independent evaluationLeaderboard updated: 2026-08-14Import ID: vals-vals_index-2026-08-16-ab641d89d6edRetrieved 2026-08-16
Vibe Code Bench v1.1Overall
70.395 %Winner
49.931 %
1.1accuracyHigher is betterCohort: vals-vibe-code-bench:1.1:overallNo documented setup differenceWinner: Gemini 3.7 FlashGemini 3.7 Flash: Harness: OpenHands
DeepSeek-V4-Pro: Harness: OpenHands
Vals AIType: Independent evaluationLeaderboard updated: 2026-08-13Import ID: vals-vibe-code-2026-08-16-a73c9770bcc4Retrieved 2026-08-16

How to read this comparison

Winning one benchmark is not an overall verdict

Your workload, budget, and required context length matter most. A coding benchmark says little about visual understanding or agent performance.

Prices use official API rates per 1M tokens. Web subscriptions and cloud platform surcharges can differ.

Every result links to its measurement source and retrieval date. Labels distinguish vendor reports, official benchmark leaderboards, and independent evaluations. Disputed or archived results are excluded.