1 article

Moonshot AI's Kimi Linear paper introduces KDA, a hybrid linear attention that beats full attention at all scales - 75% less KV cache, 6x decoding at 1M context, and open-source checkpoints.

New tutorials, open-source projects, and deep dives on coding agents - delivered weekly.
Explore 767 topics
Browse All Topics