Skip to main content
Image modelNo status listed

FLUX.2 Max

Black Forest Labs

Released
December 2025
Data date
August 15, 2026

Category view

Position within the category

This overview uses only published data from matching cohorts. Missing values never change a rank.

This position gives equal weight to 4 fixed Gradually image tests. One archived first output contributes for each model and task.

Position
Rank 7 of 28
Index score
63 / 100
Coverage
4 / 4

Leaderboard

1MAI-Image-2.5 Flash76.6 / 100
2GPT Image 271 / 100
3MAI-Image-2.567.6 / 100
6Krea 2 Large63.6 / 100
7FLUX.2 Max, Current model63 / 100
8HunyuanImage 3.0 Instruct60.5 / 100
28Stable Diffusion 3.5 Large2.5 / 100

Measurements

Comparable benchmark results

Each chart contains exactly one source, one measurement series, and one stored comparison cohort. Bars show the position. The measured value appears on the right.

Text to Image Arena

1,226 Elo · Rank 14 of 25

Sample: 9,112 · Data date: August 26, 2026

1.GPT Image 21,371 Elo
12.HiDream-O1-Image-1.51,228 Elo
13.Seedream 4.01,227 Elo
14.FLUX.2 Max, Current model1,226 Elo
15.FLUX.2 Flex1,224 Elo
16.Ideogram 4.01,223 Elo
25.FLUX.2 Klein 9B1,144 Elo

7 of 25 model versions shown in this chart. A higher value ranks first.

Image Editing Arena

1,201 Elo · Rank 13 of 19

Sample: 10,113 · Data date: August 26, 2026

1.Reve 2.11,263 Elo
11.Luma UNI 1 Max1,220 Elo
12.Nano Banana 2 Lite1,203 Elo
13.FLUX.2 Max, Current model1,201 Elo
14.HiDream-O1-Image1,190 Elo
15.Seedream 4.01,185 Elo
19.FLUX.2 Flex1,162 Elo

7 of 19 model versions shown in this chart. A higher value ranks first.

GenExam Mathematics, strict

6.6% · Rank 5 of 6

Sample: 151 · Data date: August 26, 2026

1.Nano Banana 256.3%
2.Nano Banana Pro55.6%
3.GPT Image 250.3%
4.GPT Image 1.526.5%
5.FLUX.2 Max, Current model6.6%
6.Seedream 4.02.6%

6 of 6 model versions shown in this chart. A higher value ranks first.

1,000 multidisciplinary drawing tasks with reference images and fine-grained scoring points. Strict and relaxed scoring remain separate.

Source: GenExamParticipants: 6

GenExam Physics, strict

8.8% · Rank 5 of 6

Sample: 113 · Data date: August 26, 2026

1.GPT Image 279.6%
2.Nano Banana Pro75.2%
3.Nano Banana 274.3%
4.GPT Image 1.546%
5.FLUX.2 Max, Current model8.8%
6.Seedream 4.03.5%

6 of 6 model versions shown in this chart. A higher value ranks first.

1,000 multidisciplinary drawing tasks with reference images and fine-grained scoring points. Strict and relaxed scoring remain separate.

Source: GenExamParticipants: 6

GenExam Chemistry, strict

6.8% · Rank 5 of 6

Sample: 118 · Data date: August 26, 2026

1.GPT Image 269.5%
2.Nano Banana Pro60.2%
3.Nano Banana 252.5%
4.GPT Image 1.539%
5.FLUX.2 Max, Current model6.8%
6.Seedream 4.05.9%

6 of 6 model versions shown in this chart. A higher value ranks first.

1,000 multidisciplinary drawing tasks with reference images and fine-grained scoring points. Strict and relaxed scoring remain separate.

Source: GenExamParticipants: 6

GenExam Biology, strict

11.6% · Rank 6 of 6

Sample: 156 · Data date: August 26, 2026

1.GPT Image 289.1%
2.Nano Banana Pro75.6%
3.Nano Banana 266%
4.GPT Image 1.556.4%
5.Seedream 4.018.6%
6.FLUX.2 Max, Current model11.6%

6 of 6 model versions shown in this chart. A higher value ranks first.

1,000 multidisciplinary drawing tasks with reference images and fine-grained scoring points. Strict and relaxed scoring remain separate.

Source: GenExamParticipants: 6

GenExam Geography, strict

15.2% · Rank 5 of 6

Sample: 66 · Data date: August 26, 2026

1.GPT Image 284.8%
2.Nano Banana Pro75.8%
3.Nano Banana 269.7%
4.GPT Image 1.560.6%
5.FLUX.2 Max, Current model15.2%
6.Seedream 4.010.6%

6 of 6 model versions shown in this chart. A higher value ranks first.

1,000 multidisciplinary drawing tasks with reference images and fine-grained scoring points. Strict and relaxed scoring remain separate.

Source: GenExamParticipants: 6

GenExam Computer science, strict

8.8% · Rank 5 of 6

Sample: 102 · Data date: August 26, 2026

1.GPT Image 273.5%
2.Nano Banana Pro65.7%
3.Nano Banana 256.9%
4.GPT Image 1.536.3%
5.FLUX.2 Max, Current model8.8%
6.Seedream 4.06.9%

6 of 6 model versions shown in this chart. A higher value ranks first.

1,000 multidisciplinary drawing tasks with reference images and fine-grained scoring points. Strict and relaxed scoring remain separate.

Source: GenExamParticipants: 6

GenExam Engineering, strict

10.8% · Rank 6 of 6

Sample: 111 · Data date: August 26, 2026

1.GPT Image 279.3%
2.Nano Banana Pro71.2%
3.Nano Banana 267.6%
4.GPT Image 1.544.1%
5.Seedream 4.011.7%
6.FLUX.2 Max, Current model10.8%

6 of 6 model versions shown in this chart. A higher value ranks first.

1,000 multidisciplinary drawing tasks with reference images and fine-grained scoring points. Strict and relaxed scoring remain separate.

Source: GenExamParticipants: 6

GenExam Economics, strict

2.6% · Rank 6 of 6

Sample: 77 · Data date: August 26, 2026

1.Nano Banana Pro88.3%
2.GPT Image 283.1%
3.Nano Banana 263.6%
4.GPT Image 1.542.9%
5.Seedream 4.05.2%
6.FLUX.2 Max, Current model2.6%

6 of 6 model versions shown in this chart. A higher value ranks first.

1,000 multidisciplinary drawing tasks with reference images and fine-grained scoring points. Strict and relaxed scoring remain separate.

Source: GenExamParticipants: 6

GenExam Music, strict

6.2% · Rank 5 of 6

Sample: 65 · Data date: August 26, 2026

1.GPT Image 264.6%
2.Nano Banana Pro61.5%
3.Nano Banana 250.8%
4.GPT Image 1.529.2%
5.FLUX.2 Max, Current model6.2%
6.Seedream 4.00%

6 of 6 model versions shown in this chart. A higher value ranks first.

1,000 multidisciplinary drawing tasks with reference images and fine-grained scoring points. Strict and relaxed scoring remain separate.

Source: GenExamParticipants: 6

GenExam History, strict

7.3% · Rank 5 of 6

Sample: 41 · Data date: August 26, 2026

1.Nano Banana Pro97.6%
2.GPT Image 282.9%
2.Nano Banana 282.9%
4.GPT Image 1.551.2%
5.FLUX.2 Max, Current model7.3%
5.Seedream 4.07.3%

6 of 6 model versions shown in this chart. A higher value ranks first.

1,000 multidisciplinary drawing tasks with reference images and fine-grained scoring points. Strict and relaxed scoring remain separate.

Source: GenExamParticipants: 6

GenExam overall, strict

8.5% · Rank 5 of 6

Sample: 1,000 · Data date: August 26, 2026

1.GPT Image 274.6%
2.Nano Banana Pro72.7%
3.Nano Banana 264.1%
4.GPT Image 1.543.2%
5.FLUX.2 Max, Current model8.5%
6.Seedream 4.07.2%

6 of 6 model versions shown in this chart. A higher value ranks first.

1,000 multidisciplinary drawing tasks with reference images and fine-grained scoring points. Strict and relaxed scoring remain separate.

Source: GenExamParticipants: 6

GenExam Mathematics, relaxed

49.1% · Rank 5 of 6

Sample: 151 · Data date: August 26, 2026

1.Nano Banana 287.8%
2.Nano Banana Pro86.3%
3.GPT Image 285.2%
4.GPT Image 1.565.8%
5.FLUX.2 Max, Current model49.1%
6.Seedream 4.039.8%

6 of 6 model versions shown in this chart. A higher value ranks first.

1,000 multidisciplinary drawing tasks with reference images and fine-grained scoring points. Strict and relaxed scoring remain separate.

Source: GenExamParticipants: 6

GenExam Physics, relaxed

63.2% · Rank 5 of 6

Sample: 113 · Data date: August 26, 2026

1.Nano Banana 295.7%
2.GPT Image 295.6%
3.Nano Banana Pro95.1%
4.GPT Image 1.585.4%
5.FLUX.2 Max, Current model63.2%
6.Seedream 4.049%

6 of 6 model versions shown in this chart. A higher value ranks first.

1,000 multidisciplinary drawing tasks with reference images and fine-grained scoring points. Strict and relaxed scoring remain separate.

Source: GenExamParticipants: 6

GenExam Chemistry, relaxed

54% · Rank 5 of 6

Sample: 118 · Data date: August 26, 2026

1.GPT Image 292%
2.Nano Banana 290%
3.Nano Banana Pro88.7%
4.GPT Image 1.578.1%
5.FLUX.2 Max, Current model54%
6.Seedream 4.046.1%

6 of 6 model versions shown in this chart. A higher value ranks first.

1,000 multidisciplinary drawing tasks with reference images and fine-grained scoring points. Strict and relaxed scoring remain separate.

Source: GenExamParticipants: 6

GenExam Biology, relaxed

74.5% · Rank 5 of 6

Sample: 156 · Data date: August 26, 2026

1.GPT Image 297.5%
2.Nano Banana Pro95.9%
3.Nano Banana 295.2%
4.GPT Image 1.591.9%
5.FLUX.2 Max, Current model74.5%
6.Seedream 4.071%

6 of 6 model versions shown in this chart. A higher value ranks first.

1,000 multidisciplinary drawing tasks with reference images and fine-grained scoring points. Strict and relaxed scoring remain separate.

Source: GenExamParticipants: 6

GenExam Geography, relaxed

76.3% · Rank 5 of 6

Sample: 66 · Data date: August 26, 2026

1.GPT Image 297.6%
2.Nano Banana Pro96.5%
3.Nano Banana 294.8%
4.GPT Image 1.592.5%
5.FLUX.2 Max, Current model76.3%
6.Seedream 4.065.1%

6 of 6 model versions shown in this chart. A higher value ranks first.

1,000 multidisciplinary drawing tasks with reference images and fine-grained scoring points. Strict and relaxed scoring remain separate.

Source: GenExamParticipants: 6

GenExam Computer science, relaxed

56.5% · Rank 5 of 6

Sample: 102 · Data date: August 26, 2026

1.GPT Image 293.3%
2.Nano Banana Pro91.7%
3.Nano Banana 288.8%
4.GPT Image 1.575.8%
5.FLUX.2 Max, Current model56.5%
6.Seedream 4.052.2%

6 of 6 model versions shown in this chart. A higher value ranks first.

1,000 multidisciplinary drawing tasks with reference images and fine-grained scoring points. Strict and relaxed scoring remain separate.

Source: GenExamParticipants: 6

GenExam Engineering, relaxed

68.9% · Rank 5 of 6

Sample: 111 · Data date: August 26, 2026

1.GPT Image 296.5%
2.Nano Banana 295.8%
3.Nano Banana Pro95.1%
4.GPT Image 1.586.4%
5.FLUX.2 Max, Current model68.9%
6.Seedream 4.060%

6 of 6 model versions shown in this chart. A higher value ranks first.

1,000 multidisciplinary drawing tasks with reference images and fine-grained scoring points. Strict and relaxed scoring remain separate.

Source: GenExamParticipants: 6

GenExam Economics, relaxed

61.5% · Rank 5 of 6

Sample: 77 · Data date: August 26, 2026

1.GPT Image 297.7%
2.Nano Banana Pro97.2%
3.Nano Banana 294.2%
4.GPT Image 1.585.5%
5.FLUX.2 Max, Current model61.5%
6.Seedream 4.056%

6 of 6 model versions shown in this chart. A higher value ranks first.

1,000 multidisciplinary drawing tasks with reference images and fine-grained scoring points. Strict and relaxed scoring remain separate.

Source: GenExamParticipants: 6

GenExam Music, relaxed

47% · Rank 5 of 6

Sample: 65 · Data date: August 26, 2026

1.Nano Banana Pro91%
2.GPT Image 289.1%
3.Nano Banana 286.9%
4.GPT Image 1.570.8%
5.FLUX.2 Max, Current model47%
6.Seedream 4.034.5%

6 of 6 model versions shown in this chart. A higher value ranks first.

1,000 multidisciplinary drawing tasks with reference images and fine-grained scoring points. Strict and relaxed scoring remain separate.

Source: GenExamParticipants: 6

GenExam History, relaxed

68% · Rank 5 of 6

Sample: 41 · Data date: August 26, 2026

1.Nano Banana Pro99.9%
2.Nano Banana 297.3%
3.GPT Image 297.1%
4.GPT Image 1.590.9%
5.FLUX.2 Max, Current model68%
6.Seedream 4.056.7%

6 of 6 model versions shown in this chart. A higher value ranks first.

1,000 multidisciplinary drawing tasks with reference images and fine-grained scoring points. Strict and relaxed scoring remain separate.

Source: GenExamParticipants: 6

GenExam overall, relaxed

61.9% · Rank 5 of 6

Sample: 1,000 · Data date: August 26, 2026

1.GPT Image 293.8%
2.Nano Banana Pro93.7%
3.Nano Banana 292.6%
4.GPT Image 1.582.3%
5.FLUX.2 Max, Current model61.9%
6.Seedream 4.053%

6 of 6 model versions shown in this chart. A higher value ranks first.

1,000 multidisciplinary drawing tasks with reference images and fine-grained scoring points. Strict and relaxed scoring remain separate.

Source: GenExamParticipants: 6

GRADE reasoning

47.8 points · Rank 5 of 7

Sample: 520 · Data date: August 26, 2026

1.GPT Image 282.2 points
2.Nano Banana Pro77.5 points
3.Nano Banana 272.6 points
4.GPT Image 1.554.5 points
5.FLUX.2 Max, Current model47.8 points
6.FLUX.2 Pro38.9 points
7.Seedream 4.032.4 points

7 of 7 model versions shown in this chart. A higher value ranks first.

520 scientific image-editing tasks across ten academic domains. GRADE scores reasoning, consistency, readability, and domain accuracy separately.

Source: GRADEParticipants: 7

GRADE consistency

67.2 points · Rank 5 of 7

Sample: 520 · Data date: August 26, 2026

1.GPT Image 294.4 points
2.Nano Banana Pro89.5 points
3.Nano Banana 286.4 points
4.GPT Image 1.582.3 points
5.FLUX.2 Max, Current model67.2 points
6.FLUX.2 Pro55.5 points
7.Seedream 4.053.2 points

7 of 7 model versions shown in this chart. A higher value ranks first.

520 scientific image-editing tasks across ten academic domains. GRADE scores reasoning, consistency, readability, and domain accuracy separately.

Source: GRADEParticipants: 7

GRADE readability

68.6 points · Rank 7 of 7

Sample: 520 · Data date: August 26, 2026

1.GPT Image 298.8 points
2.Nano Banana 295.9 points
3.Nano Banana Pro95.8 points
4.GPT Image 1.590.7 points
5.Seedream 4.077 points
6.FLUX.2 Pro70.3 points
7.FLUX.2 Max, Current model68.6 points

7 of 7 model versions shown in this chart. A higher value ranks first.

520 scientific image-editing tasks across ten academic domains. GRADE scores reasoning, consistency, readability, and domain accuracy separately.

Source: GRADEParticipants: 7

GRADE accuracy

11.9 points · Rank 5 of 7

Sample: 520 · Data date: August 26, 2026

1.GPT Image 256 points
2.Nano Banana Pro46.2 points
3.Nano Banana 239.6 points
4.GPT Image 1.516 points
5.FLUX.2 Max, Current model11.9 points
6.FLUX.2 Pro4.4 points
7.Seedream 4.03.1 points

7 of 7 model versions shown in this chart. A higher value ranks first.

520 scientific image-editing tasks across ten academic domains. GRADE scores reasoning, consistency, readability, and domain accuracy separately.

Source: GRADEParticipants: 7

ContraLabs typography, first-place wins

15.5% · Rank 3 of 3

Data date: August 26, 2026

1.Ideogram 4.047.9%
2.Nano Banana 230%
3.FLUX.2 Max, Current model15.5%

3 of 3 model versions shown in this chart. A higher value ranks first.

A blind typography test by ContraLabs. The provider's model card reports the first-place win rate from ten professional designers.

ContraLabs client-work suitability

2.49 rating out of 5 · Rank 3 of 3

Data date: August 26, 2026

1.Ideogram 4.03.55 rating out of 5
2.Nano Banana 22.84 rating out of 5
3.FLUX.2 Max, Current model2.49 rating out of 5

3 of 3 model versions shown in this chart. A higher value ranks first.

A blind typography test by ContraLabs. Ten professional designers rated on a 1-to-5 scale whether they would use the output in real client work.

Gradually sample test: Typography and layout

85.2 rank points · Rank 5 of 28

Sample: 1 · Data date: August 18, 2026

1.MAI-Image-2.593.8 rank points
3.Recraft V4.1 Utility92.6 rank points
4.Nano Banana 286.4 rank points
5.FLUX.2 Max, Current model85.2 rank points
6.Nano Banana 2 Lite81.5 rank points
7.FLUX.2 Flex76.6 rank points
28.Stable Diffusion 3.5 Large0 rank points

7 of 28 model versions shown in this chart. A higher value ranks first.

One archived, unfiltered first output per model. All 28 images are ranked in three model-blind orderings. The 0-100 value normalizes the mean rank. This remains an n = 1 sample, and an automated judge may have visual preferences of its own.

Gradually sample test: Product photography

56.8 rank points · Rank 12 of 28

Sample: 1 · Data date: August 18, 2026

1.GPT Image 295.1 rank points
10.Krea 2 Medium Turbo70.4 rank points
11.Luma UNI 1 Max65.4 rank points
12.FLUX.2 Max, Current model56.8 rank points
13.HiDream-O1-Image-1.555.6 rank points
14.Reve 2.151.9 rank points
28.Stable Diffusion 3.5 Large3.7 rank points

7 of 28 model versions shown in this chart. A higher value ranks first.

One archived, unfiltered first output per model. All 28 images are ranked in three model-blind orderings. The 0-100 value normalizes the mean rank. This remains an n = 1 sample, and an automated judge may have visual preferences of its own.

Gradually sample test: Character and detail

80.3 rank points · Rank 3 of 28

Sample: 1 · Data date: August 18, 2026

1.MAI-Image-2.5 Flash86.4 rank points
2.GPT Image 1.585.2 rank points
3.FLUX.2 Max, Current model80.3 rank points
4.HiDream-O1-Image-1.579 rank points
5.GPT Image 276.6 rank points
28.Stable Diffusion 3.5 Large3.7 rank points

6 of 28 model versions shown in this chart. A higher value ranks first.

One archived, unfiltered first output per model. All 28 images are ranked in three model-blind orderings. The 0-100 value normalizes the mean rank. This remains an n = 1 sample, and an automated judge may have visual preferences of its own.

Gradually sample test: Infographic

29.6 rank points · Rank 20 of 28

Sample: 1 · Data date: August 18, 2026

1.Reve 2.197.5 rank points
17.Recraft V4.1 Utility42 rank points
19.HunyuanImage 3.0 Instruct32.1 rank points
20.FLUX.2 Max, Current model29.6 rank points
21.FLUX.2 Flex25.9 rank points
22.FLUX.2 Pro24.7 rank points
28.Firefly Image Model 51.2 rank points

7 of 28 model versions shown in this chart. A higher value ranks first.

One archived, unfiltered first output per model. All 28 images are ranked in three model-blind orderings. The 0-100 value normalizes the mean rank. This remains an n = 1 sample, and an automated judge may have visual preferences of its own.

Profile

Specifications and access

Published information about this model. Unknown values are not estimated.

Weights
ProprietarySource
Access
API, WebSource
Maximum output
4 megapixelsSource
Benchmark configuration
MaxSource

Pricing

Published prices

Prices remain tied to their documented unit and source.

Representative price
$0.07 per imageSource
Together AI (black-forest-labs/FLUX.2-max)
$0.07 per megapixelSource

Measurements

Other published benchmarks

The stored dataset does not contain an exactly matching comparison cohort for these values.

LMArena Text to Image: 3D modeling

1,161.01 points

Rank 24 · 11,589 samples · Retrieved August 26, 2026

LMArena

LMArena Text to Image: Art

1,170.17 points

Rank 23 · 16,564 samples · Retrieved August 26, 2026

LMArena

LMArena Text to Image: Cartoon

1,169.23 points

Rank 23 · 42,916 samples · Retrieved August 26, 2026

LMArena

LMArena Text to Image: Commercial design

1,164.64 points

Rank 25 · 44,254 samples · Retrieved August 26, 2026

LMArena

LMArena Text to Image: Overall

1,161.94 points

Rank 22 · 117,464 samples · Retrieved August 26, 2026

LMArena

LMArena Text to Image: Photorealism

1,160.77 points

Rank 24 · 50,664 samples · Retrieved August 26, 2026

LMArena

LMArena Text to Image: Portraits

1,156.96 points

Rank 27 · 26,790 samples · Retrieved August 26, 2026

LMArena

LMArena Text to Image: Text rendering

1,167.16 points

Rank 23 · 40,386 samples · Retrieved August 26, 2026

LMArena

LMArena Image Editing: Multi-image editing

1,249.95 points

Rank 19 · 136,078 samples · Retrieved August 26, 2026

LMArena

LMArena Image Editing: Overall

1,261.76 points

Rank 27 · 353,119 samples · Retrieved August 26, 2026

LMArena

Head-to-head comparisons

Compare this model

Each matchup compares this model with exactly one other model from the same category.

More models

Models from the same selection

All AI models

Evidence

Primary sources and data date

Every statement links to its underlying documentation or leaderboard.

  • Black Forest LabsSource
  • Image pricingSource
  • Together AISource
  • Artificial Analysis (retrieved August 26, 2026)Source
  • LMArena (retrieved August 26, 2026)Source
  • LMArena (retrieved August 26, 2026)Source
  • GenExam (retrieved August 26, 2026)Source
  • GRADE (retrieved August 26, 2026)Source
  • ContraLabs via Ideogram 4 model card (retrieved August 26, 2026)Source
  • Gradually-Bildtest (retrieved August 18, 2026)Source