Decision modelOpen source
Intern-Decision-2B
InternLM
- Released
- -
- Data date
- October 3, 2026
Intern-Decision-2B is the middle multimodal variant in the Intern-Decision line. The checkpoint processes shared state, a schema of named questions, and optional images. It computes a typed JSON answer for each field in one forward pass, using Choice, Score, or Noul, and does not return free-form text.
Its two billion parameters come from fine-tuning Qwen3.5-2B. The local Apache-2.0 model uses the DecisionEngine and accepts up to 8,192 input tokens.
Specifications and access
| Specification | Value and source |
|---|---|
| Model class | Multimodal structured decision model, not a free-text generatorSource |
| Output schema | Choice, Score, and Noul as typed JSON answers with probabilitiesSource |
| Input | State, named questions, and optional imagesSource |
| Access | Local Hugging Face execution with the bundled DecisionEngineSource |
| Checkpoint | internlm/Intern-Decision-2BSource |
| Base model | Qwen/Qwen3.5-2BSource |
| Context window | 8,192-token runtime input limitSource |
| License | Apache-2.0 with the Qwen notices retainedSource |
Page 1 of 2
Measurements without matching peer values
These measurements have no matching peer values under the same test conditions. Their original values and sources remain available here.
Intern Decision evaluation: Jevbench-Easy (Show measurement, test conditions, and source)
- Source value
- 100
- Score
- 100
- Metric
- Jevbench-Easy
- Unit
- %
- Category
- Intern Decision evaluation
- Direction
- Higher is better
- Source type
- vendor-reported
- Evaluator
- InternLM Intern-Decision-2B model card
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. Local Hugging Face inference=yes; Structured candidate scoring=yes
Intern Decision evaluation: Jevbench-Original (Show measurement, test conditions, and source)
- Source value
- 84.72
- Score
- 84.72
- Metric
- Jevbench-Original
- Unit
- %
- Category
- Intern Decision evaluation
- Direction
- Higher is better
- Source type
- vendor-reported
- Evaluator
- InternLM Intern-Decision-2B model card
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. Local Hugging Face inference=yes; Structured candidate scoring=yes
Intern Decision evaluation: Jevbench-Hard (Show measurement, test conditions, and source)
- Source value
- 63.96
- Score
- 63.96
- Metric
- Jevbench-Hard
- Unit
- %
- Category
- Intern Decision evaluation
- Direction
- Higher is better
- Source type
- vendor-reported
- Evaluator
- InternLM Intern-Decision-2B model card
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. Local Hugging Face inference=yes; Structured candidate scoring=yes
Intern Decision evaluation: Typed Decision (Show measurement, test conditions, and source)
- Source value
- 79.35
- Score
- 79.35
- Metric
- Typed Decision
- Unit
- %
- Category
- Intern Decision evaluation
- Direction
- Higher is better
- Source type
- vendor-reported
- Evaluator
- InternLM Intern-Decision-2B model card
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. Local Hugging Face inference=yes; Structured candidate scoring=yes
Intern Decision evaluation: ToolACE (Show measurement, test conditions, and source)
- Source value
- 96.45
- Score
- 96.45
- Metric
- ToolACE
- Unit
- %
- Category
- Intern Decision evaluation
- Direction
- Higher is better
- Source type
- vendor-reported
- Evaluator
- InternLM Intern-Decision-2B model card
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. Local Hugging Face inference=yes; Structured candidate scoring=yes
Intern Decision evaluation: AG News (Show measurement, test conditions, and source)
- Source value
- 89.96
- Score
- 89.96
- Metric
- AG News
- Unit
- %
- Category
- Intern Decision evaluation
- Direction
- Higher is better
- Source type
- vendor-reported
- Evaluator
- InternLM Intern-Decision-2B model card
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. Local Hugging Face inference=yes; Structured candidate scoring=yes
Intern Decision evaluation: WildJailBreak (Show measurement, test conditions, and source)
- Source value
- 78.33
- Score
- 78.33
- Metric
- WildJailBreak
- Unit
- %
- Category
- Intern Decision evaluation
- Direction
- Higher is better
- Source type
- vendor-reported
- Evaluator
- InternLM Intern-Decision-2B model card
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. Local Hugging Face inference=yes; Structured candidate scoring=yes
Intern Decision evaluation: Average (Show measurement, test conditions, and source)
- Source value
- 84.68
- Score
- 84.68
- Metric
- Average
- Unit
- %
- Category
- Intern Decision evaluation
- Direction
- Higher is better
- Source type
- vendor-reported
- Evaluator
- InternLM Intern-Decision-2B model card
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. Local Hugging Face inference=yes; Structured candidate scoring=yes
Intern Decision evaluation: Brier (Show measurement, test conditions, and source)
- Source value
- 0.437
- Score
- 0.437
- Metric
- Brier
- Unit
- ratio
- Category
- Intern Decision evaluation
- Direction
- Lower is better
- Source type
- vendor-reported
- Evaluator
- InternLM Intern-Decision-2B model card
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. Local Hugging Face inference=yes; Structured candidate scoring=yes
Intern Decision evaluation: ECE (Show measurement, test conditions, and source)
- Source value
- 0.1
- Score
- 0.1
- Metric
- ECE
- Unit
- ratio
- Category
- Intern Decision evaluation
- Direction
- Lower is better
- Source type
- vendor-reported
- Evaluator
- InternLM Intern-Decision-2B model card
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. Local Hugging Face inference=yes; Structured candidate scoring=yes
Intern Decision inference latency: Mean (seconds) (Show measurement, test conditions, and source)
- Source value
- 0.03328
- Score
- 0.03328
- Metric
- Mean
- Unit
- seconds
- Category
- Intern Decision inference latency
- Direction
- Lower is better
- Source type
- vendor-reported
- Evaluator
- InternLM Intern-Decision-2B model card
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. Hardware=Single RTX 4090; Inference path=local HF; Time definition=Per-query end-to-end latency
Intern Decision inference latency: Median / P50 (seconds) (Show measurement, test conditions, and source)
- Source value
- 0.03315
- Score
- 0.03315
- Metric
- Median / P50
- Unit
- seconds
- Category
- Intern Decision inference latency
- Direction
- Lower is better
- Source type
- vendor-reported
- Evaluator
- InternLM Intern-Decision-2B model card
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. Hardware=Single RTX 4090; Inference path=local HF; Time definition=Per-query end-to-end latency
Page 1 of 2
Sources and data date
Every statement links to its underlying documentation or leaderboard.
| Type | Evidence and data date |
|---|---|
| Research status | Research date October 4, 2026. 1 source URLs checked. This documents the inspected sources, not an exhaustive inventory of every publication. |
| Additional source | Intern-Decision collection (retrieved October 3, 2026) · Intern-Decision collection · Editorial description reviewed October 3, 2026 |
| Additional source | Intern-Decision repository (retrieved October 3, 2026) · Intern-Decision repository · Editorial description reviewed October 3, 2026 |
| Additional source | Intern-Decision-2B model card (retrieved October 3, 2026; October 4, 2026) · Intern-Decision-2B model card · Editorial description reviewed October 3, 2026 |