Kimi Linear

by Moonshot AI · Kimi family · released Oct 30, 2025
Open weights Active

Kimi Linear is a 48B-total/3B-active hybrid linear attention model using Kimi Delta Attention, supporting up to 1M-token contexts with higher throughput and better long-context recall than full-attention baselines.

No pricing recorded for this model yet.

Specifications

Model typeText LLM
ArchitectureHybrid
Parameters48 (3 active)
Context window1M tokens
Max outputNot disclosed tokens
Knowledge cutoffNot disclosed
LicenseModified MIT
Modalities Text input Text output
Tags Long context Small LM
End of lifeNot disclosed
Last verifiedNot disclosed