» Tag
linear-attention
2 postsKimi Linear: Hybrid Attention Architecture Beats Full Attention
Kimi Linear's KDA module outperforms full attention, cutting KV cache by 75% and boosting throughput 6x at 1M context.
Kimi Delta Attention: A New Linear Attention Approach
Kimi Delta Attention is the latest innovation in linear attention mechanisms. Explore its important mathematical foundations for engineers.
CommitBrief — AI code reviews, right in your terminal
A provider-agnostic, local-first CLI that reviews your staged changes, a historic range, or a whole GitHub pull request. Zero telemetry, no server. Free and open source.
commitbrief.com