Language modelOpen weights
IQuest-Q1
IQuest
- Released
- -
- Data date
- October 3, 2026
IQuest-Q1 targets coding with successive tool-use steps in its post-training. The text-based MoE checkpoint handles 524,288 context tokens. Its model card estimates about 320 billion total parameters and 15 billion active parameters.
IQuest publishes local weights under its custom IQuest-Q1 license and documents SGLang and vLLM serving. Multimodal capability is not documented for this checkpoint.
Specifications and access
| Specification | Value and source |
|---|---|
| Model ID | IQuestLab/IQuest-Q1Source |
| Model class | Mixture-of-experts language model for agentic tasksSource |
| Parameters | About 320 billion total, about 15 billion activeSource |
| Context window | 524,288 tokensSource |
| Architecture | Mixture of experts with 256 total and 8 active expertsSource |
| Access | Local weights with SGLang or vLLM deploymentSource |
| License | IQuest-Q1 licenseSource |
Measurements without matching peer values
These measurements have no matching peer values under the same test conditions. Their original values and sources remain available here.
DeepSWE v1.1 (Show measurement, test conditions, and source)
- Source value
- 64.6
- Score
- 64.6
- Metric
- DeepSWE v1.1 reported score
- Unit
- No unit provided
- Category
- source-specific
- Direction
- Higher is better
- Source type
- vendor-reported
- Evaluator
- Provider model card
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. Official image-only card, visually inspected. Claude Code/Codex/mini-SWE and task-specific time limits; model card treats IQuest CLIBench as in-house.
- Context
- Source-specific observation; it is not a shared comparison cohort. The source table does not state a unit or scale for this numeric score.
NL2Repo (Show measurement, test conditions, and source)
- Source value
- 63
- Score
- 63
- Metric
- NL2Repo reported score
- Unit
- No unit provided
- Category
- source-specific
- Direction
- Higher is better
- Source type
- vendor-reported
- Evaluator
- Provider model card
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. Official image-only card, visually inspected. Claude Code/Codex/mini-SWE and task-specific time limits; model card treats IQuest CLIBench as in-house.
- Context
- Source-specific observation; it is not a shared comparison cohort. The source table does not state a unit or scale for this numeric score.
CyberGym (Show measurement, test conditions, and source)
- Source value
- 84.5
- Score
- 84.5
- Metric
- CyberGym reported score
- Unit
- No unit provided
- Category
- source-specific
- Direction
- Higher is better
- Source type
- vendor-reported
- Evaluator
- Provider model card
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. Official image-only card, visually inspected. Claude Code/Codex/mini-SWE and task-specific time limits; model card treats IQuest CLIBench as in-house.
- Context
- Source-specific observation; it is not a shared comparison cohort. The source table does not state a unit or scale for this numeric score.
TerminalBench 2.1 (Show measurement, test conditions, and source)
- Source value
- 83.2
- Score
- 83.2
- Metric
- TerminalBench 2.1 reported score
- Unit
- No unit provided
- Category
- source-specific
- Direction
- Higher is better
- Source type
- vendor-reported
- Evaluator
- Provider model card
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. Official image-only card, visually inspected. Claude Code/Codex/mini-SWE and task-specific time limits; model card treats IQuest CLIBench as in-house.
- Context
- Source-specific observation; it is not a shared comparison cohort. The source table does not state a unit or scale for this numeric score.
JobBench (Show measurement, test conditions, and source)
- Source value
- 55.7
- Score
- 55.7
- Metric
- JobBench reported score
- Unit
- No unit provided
- Category
- source-specific
- Direction
- Higher is better
- Source type
- vendor-reported
- Evaluator
- Provider model card
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. Official image-only card, visually inspected. Claude Code/Codex/mini-SWE and task-specific time limits; model card treats IQuest CLIBench as in-house.
- Context
- Source-specific observation; it is not a shared comparison cohort. The source table does not state a unit or scale for this numeric score.
Agents Last Exam (Show measurement, test conditions, and source)
- Source value
- 29.6
- Score
- 29.6
- Metric
- Agents Last Exam reported score
- Unit
- No unit provided
- Category
- source-specific
- Direction
- Higher is better
- Source type
- vendor-reported
- Evaluator
- Provider model card
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. Official image-only card, visually inspected. Claude Code/Codex/mini-SWE and task-specific time limits; model card treats IQuest CLIBench as in-house.
- Context
- Source-specific observation; it is not a shared comparison cohort. The source table does not state a unit or scale for this numeric score.
HLE no-tools (Show measurement, test conditions, and source)
- Source value
- 39.2
- Score
- 39.2
- Metric
- HLE no-tools reported score
- Unit
- No unit provided
- Category
- source-specific
- Direction
- Higher is better
- Source type
- vendor-reported
- Evaluator
- Provider model card
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. Official image-only card, visually inspected. Claude Code/Codex/mini-SWE and task-specific time limits; model card treats IQuest CLIBench as in-house.
- Context
- Source-specific observation; it is not a shared comparison cohort. The source table does not state a unit or scale for this numeric score.
IQuest CLIBench (Show measurement, test conditions, and source)
- Source value
- 53.7
- Score
- 53.7
- Metric
- IQuest CLIBench reported score
- Unit
- No unit provided
- Category
- source-specific
- Direction
- Higher is better
- Source type
- vendor-reported
- Evaluator
- Provider model card
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. Official image-only card, visually inspected. Claude Code/Codex/mini-SWE and task-specific time limits; model card treats IQuest CLIBench as in-house.
- Context
- Source-specific observation; it is not a shared comparison cohort. The source table does not state a unit or scale for this numeric score.
Sources and data date
Every statement links to its underlying documentation or leaderboard.
| Type | Evidence and data date |
|---|---|
| Research status | Research date October 4, 2026. 1 source URLs checked. This documents the inspected sources, not an exhaustive inventory of every publication. |
| Additional source | IQuest-Q1 model card (retrieved October 3, 2026; October 4, 2026) · IQuest-Q1 model card · Editorial description reviewed October 4, 2026 |