Lev
Interfaze AI
- Released
- -
- Data date
- October 3, 2026
Lev is Interfaze AI’s LoRA adapter for typed decisions over a state. You provide text, a ticket, an email, or JSON together with yes-or-no, Choice, and Score questions. Lev reads the answers from already computed logits and returns calibrated probabilities with zero output tokens. Its interface is compatible with TypeSafe’s /v1/systemone protocol.
The adapter sits on Qwen3.5-4B and is released under Apache-2.0. Separately loaded base weights complement the adapter. Lev is therefore not a standalone 4B weight checkpoint.
Specifications and access
| Specification | Value and source |
|---|---|
| Model class | System One LoRA adapter for typed decisions, not standalone base weightsSource |
| Output schema | Noul, Choice, and Score with calibrated probabilities and no output tokensSource |
| Input | Text, tickets, emails, or JSON as stateSource |
| Access | Local execution on your GPU or CPU with an optional Jev-compatible serverSource |
| Checkpoint | interfaze-ai/levSource |
| Base model | Qwen/Qwen3.5-4B, adapted through a LoRA adapterSource |
| Weights | Apache-2.0 adapter, base weights must be loaded separatelySource |
| License | Apache-2.0Source |
Page 1 of 2
Measurements without matching peer values
These measurements have no matching peer values under the same test conditions. Their original values and sources remain available here.
S1Bench: vitaminc-dev: accuracy (fraction) (Show measurement, test conditions, and source)
- Source value
- 0.668
- Score
- 0.668
- Metric
- accuracy
- Unit
- ratio
- Category
- S1Bench: vitaminc-dev
- Direction
- Higher is better
- Sample
- 3880
- Source type
- vendor-reported
- Evaluator
- Interfaze AI Lev model card
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. Total scored items=3880; Subsets=13; Harness=Nimble-pinned manifests; Checkpoint=Qwen3.5-4B LoRA adapter with provider serving runtime
S1Bench: massive-en-US: accuracy (fraction) (Show measurement, test conditions, and source)
- Source value
- 0.857
- Score
- 0.857
- Metric
- accuracy
- Unit
- ratio
- Category
- S1Bench: massive-en-US
- Direction
- Higher is better
- Sample
- 3880
- Source type
- vendor-reported
- Evaluator
- Interfaze AI Lev model card
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. Total scored items=3880; Subsets=13; Harness=Nimble-pinned manifests; Checkpoint=Qwen3.5-4B LoRA adapter with provider serving runtime
S1Bench: massive-de-DE: accuracy (fraction) (Show measurement, test conditions, and source)
- Source value
- 0.823
- Score
- 0.823
- Metric
- accuracy
- Unit
- ratio
- Category
- S1Bench: massive-de-DE
- Direction
- Higher is better
- Sample
- 3880
- Source type
- vendor-reported
- Evaluator
- Interfaze AI Lev model card
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. Total scored items=3880; Subsets=13; Harness=Nimble-pinned manifests; Checkpoint=Qwen3.5-4B LoRA adapter with provider serving runtime
S1Bench: boolq: accuracy (fraction) (Show measurement, test conditions, and source)
- Source value
- 0.827
- Score
- 0.827
- Metric
- accuracy
- Unit
- ratio
- Category
- S1Bench: boolq
- Direction
- Higher is better
- Sample
- 3880
- Source type
- vendor-reported
- Evaluator
- Interfaze AI Lev model card
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. Total scored items=3880; Subsets=13; Harness=Nimble-pinned manifests; Checkpoint=Qwen3.5-4B LoRA adapter with provider serving runtime
S1Bench: squad2: accuracy (fraction) (Show measurement, test conditions, and source)
- Source value
- 0.813
- Score
- 0.813
- Metric
- accuracy
- Unit
- ratio
- Category
- S1Bench: squad2
- Direction
- Higher is better
- Sample
- 3880
- Source type
- vendor-reported
- Evaluator
- Interfaze AI Lev model card
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. Total scored items=3880; Subsets=13; Harness=Nimble-pinned manifests; Checkpoint=Qwen3.5-4B LoRA adapter with provider serving runtime
S1Bench: paws: accuracy (fraction) (Show measurement, test conditions, and source)
- Source value
- 0.776
- Score
- 0.776
- Metric
- accuracy
- Unit
- ratio
- Category
- S1Bench: paws
- Direction
- Higher is better
- Sample
- 3880
- Source type
- vendor-reported
- Evaluator
- Interfaze AI Lev model card
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. Total scored items=3880; Subsets=13; Harness=Nimble-pinned manifests; Checkpoint=Qwen3.5-4B LoRA adapter with provider serving runtime
S1Bench: multinli: accuracy (fraction) (Show measurement, test conditions, and source)
- Source value
- 0.89
- Score
- 0.89
- Metric
- accuracy
- Unit
- ratio
- Category
- S1Bench: multinli
- Direction
- Higher is better
- Sample
- 3880
- Source type
- vendor-reported
- Evaluator
- Interfaze AI Lev model card
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. Total scored items=3880; Subsets=13; Harness=Nimble-pinned manifests; Checkpoint=Qwen3.5-4B LoRA adapter with provider serving runtime
S1Bench: civil_comments: accuracy (fraction) (Show measurement, test conditions, and source)
- Source value
- 0.76
- Score
- 0.76
- Metric
- accuracy
- Unit
- ratio
- Category
- S1Bench: civil_comments
- Direction
- Higher is better
- Sample
- 3880
- Source type
- vendor-reported
- Evaluator
- Interfaze AI Lev model card
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. Total scored items=3880; Subsets=13; Harness=Nimble-pinned manifests; Checkpoint=Qwen3.5-4B LoRA adapter with provider serving runtime
S1Bench: aegis2: accuracy (fraction) (Show measurement, test conditions, and source)
- Source value
- 0.8
- Score
- 0.8
- Metric
- accuracy
- Unit
- ratio
- Category
- S1Bench: aegis2
- Direction
- Higher is better
- Sample
- 3880
- Source type
- vendor-reported
- Evaluator
- Interfaze AI Lev model card
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. Total scored items=3880; Subsets=13; Harness=Nimble-pinned manifests; Checkpoint=Qwen3.5-4B LoRA adapter with provider serving runtime
S1Bench: helpsteer2: accuracy (fraction) (Show measurement, test conditions, and source)
- Source value
- 0.386
- Score
- 0.386
- Metric
- accuracy
- Unit
- ratio
- Category
- S1Bench: helpsteer2
- Direction
- Higher is better
- Sample
- 3880
- Source type
- vendor-reported
- Evaluator
- Interfaze AI Lev model card
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. Total scored items=3880; Subsets=13; Harness=Nimble-pinned manifests; Checkpoint=Qwen3.5-4B LoRA adapter with provider serving runtime
S1Bench: summeval-relevance: accuracy (fraction) (Show measurement, test conditions, and source)
- Source value
- 0.358
- Score
- 0.358
- Metric
- accuracy
- Unit
- ratio
- Category
- S1Bench: summeval-relevance
- Direction
- Higher is better
- Sample
- 3880
- Source type
- vendor-reported
- Evaluator
- Interfaze AI Lev model card
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. Total scored items=3880; Subsets=13; Harness=Nimble-pinned manifests; Checkpoint=Qwen3.5-4B LoRA adapter with provider serving runtime
S1Bench: summeval-consistency: accuracy (fraction) (Show measurement, test conditions, and source)
- Source value
- 0.271
- Score
- 0.271
- Metric
- accuracy
- Unit
- ratio
- Category
- S1Bench: summeval-consistency
- Direction
- Higher is better
- Sample
- 3880
- Source type
- vendor-reported
- Evaluator
- Interfaze AI Lev model card
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. Total scored items=3880; Subsets=13; Harness=Nimble-pinned manifests; Checkpoint=Qwen3.5-4B LoRA adapter with provider serving runtime
Page 1 of 2
Sources and data date
Every statement links to its underlying documentation or leaderboard.
| Type | Evidence and data date |
|---|---|
| Research status | Research date October 4, 2026. 1 source URLs checked. This documents the inspected sources, not an exhaustive inventory of every publication. |
| Unresolved | serving-revision Original wording Starred held-out values predate the last serving update; after-update rows remain separate. |
| Additional source | Lev model card (retrieved October 3, 2026; October 4, 2026) · Lev model card · Editorial description reviewed October 4, 2026 |
| Additional source | Lev announcement (retrieved October 3, 2026) · Lev announcement · Editorial description reviewed October 4, 2026 |