Skip to main content
Language modelOpen source

Ling-3.0-tiny

InclusionAI

Released
-
Data date
October 3, 2026

Ling-3.0-tiny brings InclusionAI’s hybrid attention architecture to a 7.9B checkpoint with 1.3 billion active parameters. Three Kimi Delta Attention layers alternate with one Multi-Head Latent Attention layer. The provider documents local tests on DGX Spark and Apple Silicon Macs.

InclusionAI releases BF16, FP8, and INT4 weights under MIT. The model handles 262,144 context tokens and is also available through the OpenRouter route linked in its model card.

Specifications and access

SpecificationValue and source
Parameters
7.9 billionSource
Active parameters
1.3 billion per tokenSource
Context window
262,144 tokensSource
Weights
BF16, FP8, and INT4Source
License
Access
Local weights; the model card links inclusionai/ling-3.0-tiny:free on OpenRouterSource