Kimi Linear
AssessPlatforms
The earlier Kimi model architecture that Kimi K3 scales up from.
Why it's here
Placed in Assess: 2 article(s) of evidence from 1 source(s), led by model releases, with 2 in the last 30 days. Confidence 30%. Low accumulated evidence, so it defaults conservatively pending more signal.
Evidence (2)
- 8Hacker News·7/28/2026model_releaseKimi K3 Architecture Highlights LatentMoE and NoPE
Kimi K3 is presented as a scaled-up production version of Kimi Linear, growing from 48B to 2.8T parameters and positioning itself as the largest open-weight model to date. The architecture emphasizes inference efficiency with components such as LatentMoE, Kimi Delta Attention, attention residuals, and a full switch to NoPE, while also adding native multimodal support.
- 7Hacker News·7/28/2026researchKimi Linear: Expressive Efficient Attention Architecture
Kimi Linear is a research paper describing a new attention architecture designed to improve efficiency while preserving expressive capability in large language models. The item is notable as an AI/ML research contribution rather than a product release, and it drew substantial discussion on Hacker News.