LMArena Text to Image: 3D modeling
1,258.7 points
Rank 3 · 2,907 samples · Retrieved August 26, 2026
LMArenaCategory view
This overview uses only published data from matching cohorts. Missing values never change a rank.
This position gives equal weight to 4 fixed Gradually image tests. One archived first output contributes for each model and task.
Measurements
Each chart contains exactly one source, one measurement series, and one stored comparison cohort. Bars show the position. The measured value appears on the right.
1,321 Elo · Rank 3 of 25
Sample: 15,335 · Data date: August 26, 2026
6 of 25 model versions shown in this chart. A higher value ranks first.
1,250 Elo · Rank 5 of 19
Sample: 12,690 · Data date: August 26, 2026
7 of 19 model versions shown in this chart. A higher value ranks first.
56.3% · Rank 1 of 6
Sample: 151 · Data date: August 26, 2026
6 of 6 model versions shown in this chart. A higher value ranks first.
1,000 multidisciplinary drawing tasks with reference images and fine-grained scoring points. Strict and relaxed scoring remain separate.
74.3% · Rank 3 of 6
Sample: 113 · Data date: August 26, 2026
6 of 6 model versions shown in this chart. A higher value ranks first.
1,000 multidisciplinary drawing tasks with reference images and fine-grained scoring points. Strict and relaxed scoring remain separate.
52.5% · Rank 3 of 6
Sample: 118 · Data date: August 26, 2026
6 of 6 model versions shown in this chart. A higher value ranks first.
1,000 multidisciplinary drawing tasks with reference images and fine-grained scoring points. Strict and relaxed scoring remain separate.
66% · Rank 3 of 6
Sample: 156 · Data date: August 26, 2026
6 of 6 model versions shown in this chart. A higher value ranks first.
1,000 multidisciplinary drawing tasks with reference images and fine-grained scoring points. Strict and relaxed scoring remain separate.
69.7% · Rank 3 of 6
Sample: 66 · Data date: August 26, 2026
6 of 6 model versions shown in this chart. A higher value ranks first.
1,000 multidisciplinary drawing tasks with reference images and fine-grained scoring points. Strict and relaxed scoring remain separate.
56.9% · Rank 3 of 6
Sample: 102 · Data date: August 26, 2026
6 of 6 model versions shown in this chart. A higher value ranks first.
1,000 multidisciplinary drawing tasks with reference images and fine-grained scoring points. Strict and relaxed scoring remain separate.
67.6% · Rank 3 of 6
Sample: 111 · Data date: August 26, 2026
6 of 6 model versions shown in this chart. A higher value ranks first.
1,000 multidisciplinary drawing tasks with reference images and fine-grained scoring points. Strict and relaxed scoring remain separate.
63.6% · Rank 3 of 6
Sample: 77 · Data date: August 26, 2026
6 of 6 model versions shown in this chart. A higher value ranks first.
1,000 multidisciplinary drawing tasks with reference images and fine-grained scoring points. Strict and relaxed scoring remain separate.
50.8% · Rank 3 of 6
Sample: 65 · Data date: August 26, 2026
6 of 6 model versions shown in this chart. A higher value ranks first.
1,000 multidisciplinary drawing tasks with reference images and fine-grained scoring points. Strict and relaxed scoring remain separate.
82.9% · Rank 2 of 6
Sample: 41 · Data date: August 26, 2026
6 of 6 model versions shown in this chart. A higher value ranks first.
1,000 multidisciplinary drawing tasks with reference images and fine-grained scoring points. Strict and relaxed scoring remain separate.
64.1% · Rank 3 of 6
Sample: 1,000 · Data date: August 26, 2026
6 of 6 model versions shown in this chart. A higher value ranks first.
1,000 multidisciplinary drawing tasks with reference images and fine-grained scoring points. Strict and relaxed scoring remain separate.
87.8% · Rank 1 of 6
Sample: 151 · Data date: August 26, 2026
6 of 6 model versions shown in this chart. A higher value ranks first.
1,000 multidisciplinary drawing tasks with reference images and fine-grained scoring points. Strict and relaxed scoring remain separate.
95.7% · Rank 1 of 6
Sample: 113 · Data date: August 26, 2026
6 of 6 model versions shown in this chart. A higher value ranks first.
1,000 multidisciplinary drawing tasks with reference images and fine-grained scoring points. Strict and relaxed scoring remain separate.
90% · Rank 2 of 6
Sample: 118 · Data date: August 26, 2026
6 of 6 model versions shown in this chart. A higher value ranks first.
1,000 multidisciplinary drawing tasks with reference images and fine-grained scoring points. Strict and relaxed scoring remain separate.
95.2% · Rank 3 of 6
Sample: 156 · Data date: August 26, 2026
6 of 6 model versions shown in this chart. A higher value ranks first.
1,000 multidisciplinary drawing tasks with reference images and fine-grained scoring points. Strict and relaxed scoring remain separate.
94.8% · Rank 3 of 6
Sample: 66 · Data date: August 26, 2026
6 of 6 model versions shown in this chart. A higher value ranks first.
1,000 multidisciplinary drawing tasks with reference images and fine-grained scoring points. Strict and relaxed scoring remain separate.
88.8% · Rank 3 of 6
Sample: 102 · Data date: August 26, 2026
6 of 6 model versions shown in this chart. A higher value ranks first.
1,000 multidisciplinary drawing tasks with reference images and fine-grained scoring points. Strict and relaxed scoring remain separate.
95.8% · Rank 2 of 6
Sample: 111 · Data date: August 26, 2026
6 of 6 model versions shown in this chart. A higher value ranks first.
1,000 multidisciplinary drawing tasks with reference images and fine-grained scoring points. Strict and relaxed scoring remain separate.
94.2% · Rank 3 of 6
Sample: 77 · Data date: August 26, 2026
6 of 6 model versions shown in this chart. A higher value ranks first.
1,000 multidisciplinary drawing tasks with reference images and fine-grained scoring points. Strict and relaxed scoring remain separate.
86.9% · Rank 3 of 6
Sample: 65 · Data date: August 26, 2026
6 of 6 model versions shown in this chart. A higher value ranks first.
1,000 multidisciplinary drawing tasks with reference images and fine-grained scoring points. Strict and relaxed scoring remain separate.
97.3% · Rank 2 of 6
Sample: 41 · Data date: August 26, 2026
6 of 6 model versions shown in this chart. A higher value ranks first.
1,000 multidisciplinary drawing tasks with reference images and fine-grained scoring points. Strict and relaxed scoring remain separate.
92.6% · Rank 3 of 6
Sample: 1,000 · Data date: August 26, 2026
6 of 6 model versions shown in this chart. A higher value ranks first.
1,000 multidisciplinary drawing tasks with reference images and fine-grained scoring points. Strict and relaxed scoring remain separate.
54.77 points · Rank 4 of 5
Data date: August 26, 2026
5 of 5 model versions shown in this chart. A higher value ranks first.
1,000 prompts, five top-level dimensions, and 56 detailed facets. Published top-five scores use the same deterministic judge configuration.
61.08 points · Rank 2 of 5
Data date: August 26, 2026
5 of 5 model versions shown in this chart. A higher value ranks first.
1,000 prompts, five top-level dimensions, and 56 detailed facets. Published top-five scores use the same deterministic judge configuration.
62.4 points · Rank 2 of 5
Data date: August 26, 2026
5 of 5 model versions shown in this chart. A higher value ranks first.
1,000 prompts, five top-level dimensions, and 56 detailed facets. Published top-five scores use the same deterministic judge configuration.
54.28 points · Rank 2 of 5
Data date: August 26, 2026
5 of 5 model versions shown in this chart. A higher value ranks first.
1,000 prompts, five top-level dimensions, and 56 detailed facets. Published top-five scores use the same deterministic judge configuration.
67.05 points · Rank 2 of 5
Data date: August 26, 2026
5 of 5 model versions shown in this chart. A higher value ranks first.
1,000 prompts, five top-level dimensions, and 56 detailed facets. Published top-five scores use the same deterministic judge configuration.
59.82 points · Rank 2 of 5
Data date: August 26, 2026
5 of 5 model versions shown in this chart. A higher value ranks first.
1,000 prompts, five top-level dimensions, and 56 detailed facets. Published top-five scores use the same deterministic judge configuration.
72.6 points · Rank 3 of 7
Sample: 520 · Data date: August 26, 2026
7 of 7 model versions shown in this chart. A higher value ranks first.
520 scientific image-editing tasks across ten academic domains. GRADE scores reasoning, consistency, readability, and domain accuracy separately.
86.4 points · Rank 3 of 7
Sample: 520 · Data date: August 26, 2026
7 of 7 model versions shown in this chart. A higher value ranks first.
520 scientific image-editing tasks across ten academic domains. GRADE scores reasoning, consistency, readability, and domain accuracy separately.
95.9 points · Rank 2 of 7
Sample: 520 · Data date: August 26, 2026
7 of 7 model versions shown in this chart. A higher value ranks first.
520 scientific image-editing tasks across ten academic domains. GRADE scores reasoning, consistency, readability, and domain accuracy separately.
39.6 points · Rank 3 of 7
Sample: 520 · Data date: August 26, 2026
7 of 7 model versions shown in this chart. A higher value ranks first.
520 scientific image-editing tasks across ten academic domains. GRADE scores reasoning, consistency, readability, and domain accuracy separately.
1,073 Elo · Rank 1 of 3
Data date: August 26, 2026
3 of 3 model versions shown in this chart. A higher value ranks first.
Google's model card reports Elo scores with uncertainty intervals for text-to-image generation and image editing. These rows use standard Gemini 3.1 Flash Image without tools.
1,129 Elo · Rank 1 of 3
Data date: August 26, 2026
3 of 3 model versions shown in this chart. A higher value ranks first.
Google's model card reports Elo scores with uncertainty intervals for text-to-image generation and image editing. These rows use standard Gemini 3.1 Flash Image without tools.
1,074 Elo · Rank 2 of 3
Data date: August 26, 2026
3 of 3 model versions shown in this chart. A higher value ranks first.
Google's model card reports Elo scores with uncertainty intervals for text-to-image generation and image editing. These rows use standard Gemini 3.1 Flash Image without tools.
1,047 Elo · Rank 2 of 3
Data date: August 26, 2026
3 of 3 model versions shown in this chart. A higher value ranks first.
Google's model card reports Elo scores with uncertainty intervals for text-to-image generation and image editing. These rows use standard Gemini 3.1 Flash Image without tools.
1,049 Elo · Rank 2 of 3
Data date: August 26, 2026
3 of 3 model versions shown in this chart. A higher value ranks first.
Google's model card reports Elo scores with uncertainty intervals for text-to-image generation and image editing. These rows use standard Gemini 3.1 Flash Image without tools.
1,031 Elo · Rank 1 of 3
Data date: August 26, 2026
3 of 3 model versions shown in this chart. A higher value ranks first.
Google's model card reports Elo scores with uncertainty intervals for text-to-image generation and image editing. These rows use standard Gemini 3.1 Flash Image without tools.
1,018 Elo · Rank 2 of 3
Data date: August 26, 2026
3 of 3 model versions shown in this chart. A higher value ranks first.
Google's model card reports Elo scores with uncertainty intervals for text-to-image generation and image editing. These rows use standard Gemini 3.1 Flash Image without tools.
1,016 Elo · Rank 2 of 3
Data date: August 26, 2026
3 of 3 model versions shown in this chart. A higher value ranks first.
Google's model card reports Elo scores with uncertainty intervals for text-to-image generation and image editing. These rows use standard Gemini 3.1 Flash Image without tools.
1,031 Elo · Rank 2 of 3
Data date: August 26, 2026
3 of 3 model versions shown in this chart. A higher value ranks first.
Google's model card reports Elo scores with uncertainty intervals for text-to-image generation and image editing. These rows use standard Gemini 3.1 Flash Image without tools.
30% · Rank 2 of 3
Data date: August 26, 2026
3 of 3 model versions shown in this chart. A higher value ranks first.
A blind typography test by ContraLabs. The provider's model card reports the first-place win rate from ten professional designers.
2.84 rating out of 5 · Rank 2 of 3
Data date: August 26, 2026
3 of 3 model versions shown in this chart. A higher value ranks first.
A blind typography test by ContraLabs. Ten professional designers rated on a 1-to-5 scale whether they would use the output in real client work.
86.4 rank points · Rank 4 of 28
Sample: 1 · Data date: August 18, 2026
7 of 28 model versions shown in this chart. A higher value ranks first.
One archived, unfiltered first output per model. All 28 images are ranked in three model-blind orderings. The 0-100 value normalizes the mean rank. This remains an n = 1 sample, and an automated judge may have visual preferences of its own.
29.6 rank points · Rank 20 of 28
Sample: 1 · Data date: August 18, 2026
7 of 28 model versions shown in this chart. A higher value ranks first.
One archived, unfiltered first output per model. All 28 images are ranked in three model-blind orderings. The 0-100 value normalizes the mean rank. This remains an n = 1 sample, and an automated judge may have visual preferences of its own.
18.5 rank points · Rank 25 of 28
Sample: 1 · Data date: August 18, 2026
7 of 28 model versions shown in this chart. A higher value ranks first.
One archived, unfiltered first output per model. All 28 images are ranked in three model-blind orderings. The 0-100 value normalizes the mean rank. This remains an n = 1 sample, and an automated judge may have visual preferences of its own.
49.4 rank points · Rank 14 of 28
Sample: 1 · Data date: August 18, 2026
7 of 28 model versions shown in this chart. A higher value ranks first.
One archived, unfiltered first output per model. All 28 images are ranked in three model-blind orderings. The 0-100 value normalizes the mean rank. This remains an n = 1 sample, and an automated judge may have visual preferences of its own.
Profile
Published information about this model. Unknown values are not estimated.
Pricing
Prices remain tied to their documented unit and source.
Measurements
The stored dataset does not contain an exactly matching comparison cohort for these values.
1,258.7 points
Rank 3 · 2,907 samples · Retrieved August 26, 2026
LMArena1,250.51 points
Rank 8 · 4,154 samples · Retrieved August 26, 2026
LMArena1,266.73 points
Rank 8 · 10,445 samples · Retrieved August 26, 2026
LMArena1,272.16 points
Rank 4 · 10,316 samples · Retrieved August 26, 2026
LMArena1,263.06 points
Rank 5 · 27,910 samples · Retrieved August 26, 2026
LMArena1,269.81 points
Rank 6 · 11,229 samples · Retrieved August 26, 2026
LMArena1,266.8 points
Rank 7 · 5,554 samples · Retrieved August 26, 2026
LMArena1,294.78 points
Rank 5 · 9,489 samples · Retrieved August 26, 2026
LMArena1,364.76 points
Rank 5 · 37,577 samples · Retrieved August 26, 2026
LMArena1,386.27 points
Rank 10 · 117,943 samples · Retrieved August 26, 2026
LMArenaHead-to-head comparisons
Each matchup compares this model with exactly one other model from the same category.
More models
All AI models
Evidence
Every statement links to its underlying documentation or leaderboard.