MAI-Code-1.1-Flash
Microsoft AI
- Released
- August 11, 2026
- Data date
- October 3, 2026
MAI-Code-1.1-Flash reduces token use on coding tasks. Microsoft reports 25% fewer tokens per task and 25% faster token streaming in GitHub Copilot compared with version 1.0. Its post-training focuses in part on CLI tasks and .NET.
Microsoft put the version into production in GitHub Copilot on August 11, 2026. The release offers neither open weights nor a separate public API endpoint.
Measurements without matching peer values
These measurements have no matching peer values under the same test conditions. Their original values and sources remain available here.
SWE-bench Verified pass rate (Show measurement, test conditions, and source)
- Source value
- 72.6
- Score
- 72.6
- Metric
- SWE-bench Verified pass rate
- Unit
- %
- Category
- coding
- Direction
- Higher is better
- Source type
- vendor-reported
- Evaluator
- MAI-Code-1.1-Flash Model Card
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. GitHub Copilot production harness; VS Code repository context, tool calls, and verification; model-card page 5.
- Context
- Vendor-reported, harness-specific result. Reasoning effort, sample count, and evaluation date are unspecified; not a shared comparison cohort.
Terminal-Bench 2.1 pass rate (Show measurement, test conditions, and source)
- Source value
- 62.9
- Score
- 62.9
- Metric
- Terminal-Bench 2.1 pass rate
- Unit
- %
- Category
- coding
- Direction
- Higher is better
- Source type
- vendor-reported
- Evaluator
- MAI-Code-1.1-Flash Model Card
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. GitHub Copilot production harness; VS Code repository context, tool calls, and verification; model-card page 5.
- Context
- Vendor-reported, harness-specific result. Reasoning effort, sample count, and evaluation date are unspecified; not a shared comparison cohort.
Text2WebApp pass rate (Show measurement, test conditions, and source)
- Source value
- 74.1
- Score
- 74.1
- Metric
- Text2WebApp pass rate
- Unit
- %
- Category
- coding
- Direction
- Higher is better
- Source type
- vendor-reported
- Evaluator
- MAI-Code-1.1-Flash Model Card
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. Internal visual-coding evaluation in the GitHub Copilot production harness; text specification to web application; model-card page 5.
- Context
- Vendor-reported internal benchmark. Public tasks, grading protocol, sample count, reasoning effort, and run date are unspecified.
ScreenShot2WebApp pass rate (Show measurement, test conditions, and source)
- Source value
- 42.1
- Score
- 42.1
- Metric
- ScreenShot2WebApp pass rate
- Unit
- %
- Category
- coding
- Direction
- Higher is better
- Source type
- vendor-reported
- Evaluator
- MAI-Code-1.1-Flash Model Card
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. Internal visual-coding evaluation in the GitHub Copilot production harness; screenshot to web application; model-card page 5.
- Context
- Vendor-reported internal benchmark. Public tasks, grading protocol, sample count, reasoning effort, and run date are unspecified.
Vision2Web Level 3 pass rate (Show measurement, test conditions, and source)
- Source value
- 11.5
- Score
- 11.5
- Metric
- Vision2Web Level 3 pass rate
- Unit
- %
- Category
- coding
- Direction
- Higher is better
- Source type
- vendor-reported
- Evaluator
- MAI-Code-1.1-Flash Model Card
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. Source-specific Vision2Web Level 3 evaluation in the GitHub Copilot production harness; model-card page 5.
- Context
- Vendor-reported visual-coding benchmark. This card does not provide public tasks, grading protocol, sample count, reasoning effort, or run date, nor explicitly label Vision2Web internal.
Sources and data date
Every statement links to its underlying documentation or leaderboard.
| Type | Evidence and data date |
|---|---|
| Research status | Research date October 4, 2026. 3 source URLs checked. This documents the inspected sources, not an exhaustive inventory of every publication. |
| Not published | vision-benchmark-protocol Original wording The card reports exact pass rates for the three visual-coding tests without public tasks, grading protocol, or sample count. Only Text2WebApp and ScreenShot2WebApp are explicitly labelled internal; that label is not inferred for Vision2Web. |
| Additional source | Microsoft AI MAI-Code-1.1-Flash announcement (retrieved October 3, 2026; October 4, 2026) · Microsoft AI MAI-Code-1.1-Flash announcement · Editorial description reviewed October 4, 2026 |
| Additional source | MAI-Code-1.1-Flash Model Card (retrieved October 4, 2026) |