Language modelOpen source
Ming-flash-omni-2.0
InclusionAI
- Released
- February 11, 2026
- Data date
- October 3, 2026
Ming-flash-omni-2.0 combines image editing, video conversations, and controllable audio output in one model. InclusionAI’s omnimodal checkpoint accepts text, images, video, and audio, and generates text, images, and audio.
Its MoE architecture activates 6 of its 100 billion parameters. InclusionAI released the weights on February 11, 2026 under MIT.
Page 1 of 2
Measurements without matching peer values
These measurements have no matching peer values under the same test conditions. Their original values and sources remain available here.
MMStar (Show measurement, test conditions, and source)
- Source value
- 74.88
- Score
- 74.88
- Metric
- MMStar
- Unit
- %
- Category
- source-specific
- Direction
- Higher is better
- Source type
- vendor-reported
- Evaluator
- HF model-card benchmark image linked to technical report
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. Multimodal/audio/video/tool-use/generation comparison table; no single independent evaluator cohort should be inferred.
- Context
- Source-specific observation; it is not a shared comparison cohort. Provider card; image footnote mixes report values with marked public-API/local-deployment rows and specialist output models.
MMBench_EN_test11 (Show measurement, test conditions, and source)
- Source value
- 87
- Score
- 87
- Metric
- MMBench_EN_test11
- Unit
- %
- Category
- source-specific
- Direction
- Higher is better
- Source type
- vendor-reported
- Evaluator
- HF model-card benchmark image linked to technical report
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. Multimodal/audio/video/tool-use/generation comparison table; no single independent evaluator cohort should be inferred.
- Context
- Source-specific observation; it is not a shared comparison cohort. Provider card; image footnote mixes report values with marked public-API/local-deployment rows and specialist output models.
HallusionBench (Show measurement, test conditions, and source)
- Source value
- 66.14
- Score
- 66.14
- Metric
- HallusionBench
- Unit
- %
- Category
- source-specific
- Direction
- Higher is better
- Source type
- vendor-reported
- Evaluator
- HF model-card benchmark image linked to technical report
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. Multimodal/audio/video/tool-use/generation comparison table; no single independent evaluator cohort should be inferred.
- Context
- Source-specific observation; it is not a shared comparison cohort. Provider card; image footnote mixes report values with marked public-API/local-deployment rows and specialist output models.
MMVet (Show measurement, test conditions, and source)
- Source value
- 85.64
- Score
- 85.64
- Metric
- MMVet
- Unit
- %
- Category
- source-specific
- Direction
- Higher is better
- Source type
- vendor-reported
- Evaluator
- HF model-card benchmark image linked to technical report
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. Multimodal/audio/video/tool-use/generation comparison table; no single independent evaluator cohort should be inferred.
- Context
- Source-specific observation; it is not a shared comparison cohort. Provider card; image footnote mixes report values with marked public-API/local-deployment rows and specialist output models.
MMMU (Show measurement, test conditions, and source)
- Source value
- 77.11
- Score
- 77.11
- Metric
- MMMU
- Unit
- %
- Category
- source-specific
- Direction
- Higher is better
- Source type
- vendor-reported
- Evaluator
- HF model-card benchmark image linked to technical report
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. Multimodal/audio/video/tool-use/generation comparison table; no single independent evaluator cohort should be inferred.
- Context
- Source-specific observation; it is not a shared comparison cohort. Provider card; image footnote mixes report values with marked public-API/local-deployment rows and specialist output models.
MVbench (Show measurement, test conditions, and source)
- Source value
- 77.5
- Score
- 77.5
- Metric
- MVbench
- Unit
- %
- Category
- source-specific
- Direction
- Higher is better
- Source type
- vendor-reported
- Evaluator
- HF model-card benchmark image linked to technical report
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. Multimodal/audio/video/tool-use/generation comparison table; no single independent evaluator cohort should be inferred.
- Context
- Source-specific observation; it is not a shared comparison cohort. Provider card; image footnote mixes report values with marked public-API/local-deployment rows and specialist output models.
MLVU (Show measurement, test conditions, and source)
- Source value
- 77.72
- Score
- 77.72
- Metric
- MLVU
- Unit
- %
- Category
- source-specific
- Direction
- Higher is better
- Source type
- vendor-reported
- Evaluator
- HF model-card benchmark image linked to technical report
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. Multimodal/audio/video/tool-use/generation comparison table; no single independent evaluator cohort should be inferred.
- Context
- Source-specific observation; it is not a shared comparison cohort. Provider card; image footnote mixes report values with marked public-API/local-deployment rows and specialist output models.
librispeech_test-other (down) (Show measurement, test conditions, and source)
- Source value
- 2.2
- Score
- 2.2
- Metric
- librispeech_test-other (down)
- Unit
- %
- Category
- source-specific
- Direction
- Lower is better
- Source type
- vendor-reported
- Evaluator
- HF model-card benchmark image linked to technical report
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. Multimodal/audio/video/tool-use/generation comparison table; no single independent evaluator cohort should be inferred.
- Context
- Source-specific observation; it is not a shared comparison cohort. Provider card; image footnote mixes report values with marked public-API/local-deployment rows and specialist output models.
AlpacaEval (Show measurement, test conditions, and source)
- Source value
- 95.6
- Score
- 95.6
- Metric
- AlpacaEval
- Unit
- %
- Category
- source-specific
- Direction
- Higher is better
- Source type
- vendor-reported
- Evaluator
- HF model-card benchmark image linked to technical report
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. Multimodal/audio/video/tool-use/generation comparison table; no single independent evaluator cohort should be inferred.
- Context
- Source-specific observation; it is not a shared comparison cohort. Provider card; image footnote mixes report values with marked public-API/local-deployment rows and specialist output models.
ContextASR-Bench (down) (Show measurement, test conditions, and source)
- Source value
- 3.3
- Score
- 3.3
- Metric
- ContextASR-Bench (down)
- Unit
- %
- Category
- source-specific
- Direction
- Lower is better
- Source type
- vendor-reported
- Evaluator
- HF model-card benchmark image linked to technical report
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. Multimodal/audio/video/tool-use/generation comparison table; no single independent evaluator cohort should be inferred.
- Context
- Source-specific observation; it is not a shared comparison cohort. Provider card; image footnote mixes report values with marked public-API/local-deployment rows and specialist output models.
WSC-Eval-TTS-Avg-WER (down) (Show measurement, test conditions, and source)
- Source value
- 2.71
- Score
- 2.71
- Metric
- WSC-Eval-TTS-Avg-WER (down)
- Unit
- %
- Category
- source-specific
- Direction
- Lower is better
- Source type
- vendor-reported
- Evaluator
- HF model-card benchmark image linked to technical report
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. Multimodal/audio/video/tool-use/generation comparison table; no single independent evaluator cohort should be inferred.
- Context
- Source-specific observation; it is not a shared comparison cohort. Provider card; image footnote mixes report values with marked public-API/local-deployment rows and specialist output models.
WSC-Eval-TTS-Avg-ACC (Show measurement, test conditions, and source)
- Source value
- 83.25
- Score
- 83.25
- Metric
- WSC-Eval-TTS-Avg-ACC
- Unit
- %
- Category
- source-specific
- Direction
- Higher is better
- Source type
- vendor-reported
- Evaluator
- HF model-card benchmark image linked to technical report
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. Multimodal/audio/video/tool-use/generation comparison table; no single independent evaluator cohort should be inferred.
- Context
- Source-specific observation; it is not a shared comparison cohort. Provider card; image footnote mixes report values with marked public-API/local-deployment rows and specialist output models.
Page 1 of 2
Sources and data date
Every statement links to its underlying documentation or leaderboard.
| Type | Evidence and data date |
|---|---|
| Research status | Research date October 4, 2026. 1 source URLs checked. This documents the inspected sources, not an exhaustive inventory of every publication. |
| Additional source | Ming-flash-omni 2.0 model card (retrieved October 3, 2026; October 4, 2026) · Ming-flash-omni 2.0 model card · Editorial description reviewed October 4, 2026 |