Skip to main content
Image modelNo status listed

FLUX.2 Pro

Black Forest Labs

Released
November 2025
Data date
August 15, 2026

Category view

Position within the category

This overview uses only published data from matching cohorts. Missing values never change a rank.

This position gives equal weight to 4 fixed Gradually image tests. One archived first output contributes for each model and task.

Position
Rank 18 of 28
Index score
47.2 / 100
Coverage
4 / 4

Leaderboard

1MAI-Image-2.5 Flash76.6 / 100
2GPT Image 271 / 100
3MAI-Image-2.567.6 / 100
18FLUX.2 Flex47.2 / 100
18FLUX.2 Pro, Current model47.2 / 100
20Nano Banana 246 / 100
28Stable Diffusion 3.5 Large2.5 / 100

Measurements

Comparable benchmark results

Each chart contains exactly one source, one measurement series, and one stored comparison cohort. Bars show the position. The measured value appears on the right.

Text to Image Arena

1,208 Elo · Rank 22 of 25

Sample: 8,915 · Data date: August 26, 2026

1.GPT Image 21,371 Elo
20.Krea 2 Medium Turbo1,215 Elo
21.Wan 2.6 Text to Image1,210 Elo
22.FLUX.2 Pro, Current model1,208 Elo
23.HiDream-O1-Image1,175 Elo
24.HunyuanImage 3.0 Instruct1,150 Elo
25.FLUX.2 Klein 9B1,144 Elo

7 of 25 model versions shown in this chart. A higher value ranks first.

Image Editing Arena

1,170 Elo · Rank 16 of 19

Sample: 11,529 · Data date: August 26, 2026

1.Reve 2.11,263 Elo
14.HiDream-O1-Image1,190 Elo
15.Seedream 4.01,185 Elo
16.FLUX.2 Pro, Current model1,170 Elo
17.FLUX.2 Klein 9B1,166 Elo
18.Qwen Image 2.0 Pro1,164 Elo
19.FLUX.2 Flex1,162 Elo

7 of 19 model versions shown in this chart. A higher value ranks first.

GRADE reasoning

38.9 points · Rank 6 of 7

Sample: 520 · Data date: August 26, 2026

1.GPT Image 282.2 points
2.Nano Banana Pro77.5 points
3.Nano Banana 272.6 points
4.GPT Image 1.554.5 points
5.FLUX.2 Max47.8 points
6.FLUX.2 Pro, Current model38.9 points
7.Seedream 4.032.4 points

7 of 7 model versions shown in this chart. A higher value ranks first.

520 scientific image-editing tasks across ten academic domains. GRADE scores reasoning, consistency, readability, and domain accuracy separately.

Source: GRADEParticipants: 7

GRADE consistency

55.5 points · Rank 6 of 7

Sample: 520 · Data date: August 26, 2026

1.GPT Image 294.4 points
2.Nano Banana Pro89.5 points
3.Nano Banana 286.4 points
4.GPT Image 1.582.3 points
5.FLUX.2 Max67.2 points
6.FLUX.2 Pro, Current model55.5 points
7.Seedream 4.053.2 points

7 of 7 model versions shown in this chart. A higher value ranks first.

520 scientific image-editing tasks across ten academic domains. GRADE scores reasoning, consistency, readability, and domain accuracy separately.

Source: GRADEParticipants: 7

GRADE readability

70.3 points · Rank 6 of 7

Sample: 520 · Data date: August 26, 2026

1.GPT Image 298.8 points
2.Nano Banana 295.9 points
3.Nano Banana Pro95.8 points
4.GPT Image 1.590.7 points
5.Seedream 4.077 points
6.FLUX.2 Pro, Current model70.3 points
7.FLUX.2 Max68.6 points

7 of 7 model versions shown in this chart. A higher value ranks first.

520 scientific image-editing tasks across ten academic domains. GRADE scores reasoning, consistency, readability, and domain accuracy separately.

Source: GRADEParticipants: 7

GRADE accuracy

4.4 points · Rank 6 of 7

Sample: 520 · Data date: August 26, 2026

1.GPT Image 256 points
2.Nano Banana Pro46.2 points
3.Nano Banana 239.6 points
4.GPT Image 1.516 points
5.FLUX.2 Max11.9 points
6.FLUX.2 Pro, Current model4.4 points
7.Seedream 4.03.1 points

7 of 7 model versions shown in this chart. A higher value ranks first.

520 scientific image-editing tasks across ten academic domains. GRADE scores reasoning, consistency, readability, and domain accuracy separately.

Source: GRADEParticipants: 7

GEBench Chinese, single-step

68.83 points · Rank 3 of 5

Data date: August 26, 2026

1.Nano Banana Pro84.5 points
2.GPT Image 1.583.79 points
3.FLUX.2 Pro, Current model68.83 points
4.Wan 2.6 Text to Image64.2 points
5.Seedream 4.062.04 points

5 of 5 model versions shown in this chart. A higher value ranks first.

A bilingual graphical-user-interface benchmark. Chinese and English subsets and the single-step, multi-step, fictional-app, real-app, and grounding categories remain separate.

Source: GEBenchParticipants: 5

GEBench Chinese, multi-step

55.07 points · Rank 3 of 5

Data date: August 26, 2026

1.Nano Banana Pro68.65 points
2.GPT Image 1.556.97 points
3.FLUX.2 Pro, Current model55.07 points
4.Wan 2.6 Text to Image50.11 points
5.Seedream 4.048.64 points

5 of 5 model versions shown in this chart. A higher value ranks first.

A bilingual graphical-user-interface benchmark. Chinese and English subsets and the single-step, multi-step, fictional-app, real-app, and grounding categories remain separate.

Source: GEBenchParticipants: 5

GEBench Chinese, fictional app

58.13 points · Rank 3 of 5

Data date: August 26, 2026

1.Nano Banana Pro65.75 points
2.GPT Image 1.560.11 points
3.FLUX.2 Pro, Current model58.13 points
4.Wan 2.6 Text to Image52.72 points
5.Seedream 4.049.28 points

5 of 5 model versions shown in this chart. A higher value ranks first.

A bilingual graphical-user-interface benchmark. Chinese and English subsets and the single-step, multi-step, fictional-app, real-app, and grounding categories remain separate.

Source: GEBenchParticipants: 5

GEBench Chinese, real app

55.41 points · Rank 3 of 5

Data date: August 26, 2026

1.Nano Banana Pro64.35 points
2.GPT Image 1.555.65 points
3.FLUX.2 Pro, Current model55.41 points
4.Seedream 4.050.93 points
5.Wan 2.6 Text to Image50.4 points

5 of 5 model versions shown in this chart. A higher value ranks first.

A bilingual graphical-user-interface benchmark. Chinese and English subsets and the single-step, multi-step, fictional-app, real-app, and grounding categories remain separate.

Source: GEBenchParticipants: 5

GEBench Chinese, grounding

50.24 points · Rank 5 of 5

Data date: August 26, 2026

1.Nano Banana Pro64.83 points
2.Wan 2.6 Text to Image59.58 points
3.Seedream 4.053.53 points
4.GPT Image 1.553.33 points
5.FLUX.2 Pro, Current model50.24 points

5 of 5 model versions shown in this chart. A higher value ranks first.

A bilingual graphical-user-interface benchmark. Chinese and English subsets and the single-step, multi-step, fictional-app, real-app, and grounding categories remain separate.

Source: GEBenchParticipants: 5

GEBench Chinese, overall

57.54 points · Rank 3 of 5

Data date: August 26, 2026

1.Nano Banana Pro69.62 points
2.GPT Image 1.563.22 points
3.FLUX.2 Pro, Current model57.54 points
4.Wan 2.6 Text to Image55.4 points
5.Seedream 4.052.88 points

5 of 5 model versions shown in this chart. A higher value ranks first.

A bilingual graphical-user-interface benchmark. Chinese and English subsets and the single-step, multi-step, fictional-app, real-app, and grounding categories remain separate.

Source: GEBenchParticipants: 5

GEBench English, single-step

61 points · Rank 3 of 5

Data date: August 26, 2026

1.Nano Banana Pro84.32 points
2.GPT Image 1.580.8 points
3.FLUX.2 Pro, Current model61 points
4.Wan 2.6 Text to Image60.17 points
5.Seedream 4.053.28 points

5 of 5 model versions shown in this chart. A higher value ranks first.

A bilingual graphical-user-interface benchmark. Chinese and English subsets and the single-step, multi-step, fictional-app, real-app, and grounding categories remain separate.

Source: GEBenchParticipants: 5

GEBench English, multi-step

52.17 points · Rank 3 of 5

Data date: August 26, 2026

1.Nano Banana Pro69.51 points
2.GPT Image 1.558.87 points
3.FLUX.2 Pro, Current model52.17 points
4.Wan 2.6 Text to Image44.36 points
5.Seedream 4.037.57 points

5 of 5 model versions shown in this chart. A higher value ranks first.

A bilingual graphical-user-interface benchmark. Chinese and English subsets and the single-step, multi-step, fictional-app, real-app, and grounding categories remain separate.

Source: GEBenchParticipants: 5

GEBench English, fictional app

49.92 points · Rank 2 of 5

Data date: August 26, 2026

1.GPT Image 1.563.68 points
2.FLUX.2 Pro, Current model49.92 points
3.Wan 2.6 Text to Image49.55 points
4.Seedream 4.047.92 points
5.Nano Banana Pro46.33 points

5 of 5 model versions shown in this chart. A higher value ranks first.

A bilingual graphical-user-interface benchmark. Chinese and English subsets and the single-step, multi-step, fictional-app, real-app, and grounding categories remain separate.

Source: GEBenchParticipants: 5

GEBench English, real app

47.16 points · Rank 4 of 5

Data date: August 26, 2026

1.GPT Image 1.558.93 points
2.Seedream 4.049.36 points
3.Nano Banana Pro47.2 points
4.FLUX.2 Pro, Current model47.16 points
5.Wan 2.6 Text to Image44.8 points

5 of 5 model versions shown in this chart. A higher value ranks first.

A bilingual graphical-user-interface benchmark. Chinese and English subsets and the single-step, multi-step, fictional-app, real-app, and grounding categories remain separate.

Source: GEBenchParticipants: 5

GEBench English, grounding

45.67 points · Rank 4 of 5

Data date: August 26, 2026

1.Nano Banana Pro58.64 points
2.Wan 2.6 Text to Image53.36 points
3.GPT Image 1.549.23 points
4.FLUX.2 Pro, Current model45.67 points
5.Seedream 4.044.17 points

5 of 5 model versions shown in this chart. A higher value ranks first.

A bilingual graphical-user-interface benchmark. Chinese and English subsets and the single-step, multi-step, fictional-app, real-app, and grounding categories remain separate.

Source: GEBenchParticipants: 5

GEBench English, overall

51.18 points · Rank 3 of 5

Data date: August 26, 2026

1.GPT Image 1.563.16 points
2.Nano Banana Pro61.2 points
3.FLUX.2 Pro, Current model51.18 points
4.Wan 2.6 Text to Image50.45 points
5.Seedream 4.046.46 points

5 of 5 model versions shown in this chart. A higher value ranks first.

A bilingual graphical-user-interface benchmark. Chinese and English subsets and the single-step, multi-step, fictional-app, real-app, and grounding categories remain separate.

Source: GEBenchParticipants: 5

Gradually sample test: Typography and layout

50.6 rank points · Rank 11 of 28

Sample: 1 · Data date: August 18, 2026

1.MAI-Image-2.593.8 rank points
9.Firefly Image Model 575.3 rank points
10.Grok Imagine Image Quality65.4 rank points
11.FLUX.2 Pro, Current model50.6 rank points
11.Luma UNI 1 Max50.6 rank points
13.HiDream-O1-Image-1.549.4 rank points
28.Stable Diffusion 3.5 Large0 rank points

7 of 28 model versions shown in this chart. A higher value ranks first.

One archived, unfiltered first output per model. All 28 images are ranked in three model-blind orderings. The 0-100 value normalizes the mean rank. This remains an n = 1 sample, and an automated judge may have visual preferences of its own.

Gradually sample test: Product photography

43.2 rank points · Rank 16 of 28

Sample: 1 · Data date: August 18, 2026

1.GPT Image 295.1 rank points
14.Reve 2.151.9 rank points
15.FLUX.2 Flex46.9 rank points
16.FLUX.2 Pro, Current model43.2 rank points
17.Qwen Image 2.0 Pro42 rank points
18.Midjourney V8.140.7 rank points
28.Stable Diffusion 3.5 Large3.7 rank points

7 of 28 model versions shown in this chart. A higher value ranks first.

One archived, unfiltered first output per model. All 28 images are ranked in three model-blind orderings. The 0-100 value normalizes the mean rank. This remains an n = 1 sample, and an automated judge may have visual preferences of its own.

Gradually sample test: Character and detail

70.4 rank points · Rank 8 of 28

Sample: 1 · Data date: August 18, 2026

1.MAI-Image-2.5 Flash86.4 rank points
5.Qwen Image 2.0 Pro76.6 rank points
7.Recraft V4.1 Utility72.9 rank points
8.FLUX.2 Pro, Current model70.4 rank points
9.Firefly Image Model 566.7 rank points
9.Wan 2.6 Text to Image66.7 rank points
28.Stable Diffusion 3.5 Large3.7 rank points

7 of 28 model versions shown in this chart. A higher value ranks first.

One archived, unfiltered first output per model. All 28 images are ranked in three model-blind orderings. The 0-100 value normalizes the mean rank. This remains an n = 1 sample, and an automated judge may have visual preferences of its own.

Gradually sample test: Infographic

24.7 rank points · Rank 22 of 28

Sample: 1 · Data date: August 18, 2026

1.Reve 2.197.5 rank points
20.FLUX.2 Max29.6 rank points
21.FLUX.2 Flex25.9 rank points
22.FLUX.2 Pro, Current model24.7 rank points
23.Grok Imagine Image Quality21 rank points
24.FLUX.2 Klein 9B12.3 rank points
28.Firefly Image Model 51.2 rank points

7 of 28 model versions shown in this chart. A higher value ranks first.

One archived, unfiltered first output per model. All 28 images are ranked in three model-blind orderings. The 0-100 value normalizes the mean rank. This remains an n = 1 sample, and an automated judge may have visual preferences of its own.

Profile

Specifications and access

Published information about this model. Unknown values are not estimated.

Weights
ProprietarySource
Access
API, WebSource
Maximum output
4 megapixelsSource
Benchmark configuration
ProSource

Pricing

Published prices

Prices remain tied to their documented unit and source.

Representative price
$0.03 per imageSource
Together AI (black-forest-labs/FLUX.2-pro)
$0.03 per imageSource

Measurements

Other published benchmarks

The stored dataset does not contain an exactly matching comparison cohort for these values.

LMArena Text to Image: 3D modeling

1,149.39 points

Rank 28 · 17,241 samples · Retrieved August 26, 2026

LMArena

LMArena Text to Image: Art

1,163.05 points

Rank 26 · 25,069 samples · Retrieved August 26, 2026

LMArena

LMArena Text to Image: Cartoon

1,161.02 points

Rank 27 · 66,010 samples · Retrieved August 26, 2026

LMArena

LMArena Text to Image: Commercial design

1,156.46 points

Rank 26 · 64,392 samples · Retrieved August 26, 2026

LMArena

LMArena Text to Image: Overall

1,154.79 points

Rank 26 · 175,196 samples · Retrieved August 26, 2026

LMArena

LMArena Text to Image: Photorealism

1,151.5 points

Rank 30 · 74,515 samples · Retrieved August 26, 2026

LMArena

LMArena Text to Image: Portraits

1,146.51 points

Rank 31 · 39,262 samples · Retrieved August 26, 2026

LMArena

LMArena Text to Image: Text rendering

1,156.45 points

Rank 27 · 59,326 samples · Retrieved August 26, 2026

LMArena

LMArena Image Editing: Multi-image editing

1,238.24 points

Rank 20 · 171,464 samples · Retrieved August 26, 2026

LMArena

LMArena Image Editing: Overall

1,244.29 points

Rank 30 · 497,149 samples · Retrieved August 26, 2026

LMArena

Head-to-head comparisons

Compare this model

Each matchup compares this model with exactly one other model from the same category.

More models

Models from the same selection

All AI models

Evidence

Primary sources and data date

Every statement links to its underlying documentation or leaderboard.

  • Black Forest LabsSource
  • Image pricingSource
  • Together AISource
  • Artificial Analysis (retrieved August 26, 2026)Source
  • LMArena (retrieved August 26, 2026)Source
  • LMArena (retrieved August 26, 2026)Source
  • GRADE (retrieved August 26, 2026)Source
  • GEBench (retrieved August 26, 2026)Source
  • Gradually-Bildtest (retrieved August 18, 2026)Source