» Tag
quantization
24 postsIntel's ACE brings outer-product matrix math to x86, rivaling Arm SME
Intel's ACE extends AMX with outer-product matrix acceleration, compared against Arm's SME2 in a detailed technical breakdown.
Bonsai 27B Runs Locally on iPhone - A 27B Model in 3.9GB
Bonsai 27B model is developed with 1-bit quantization for local iPhone use.
KV Cache Quantization's Effect on KLD in Qwen3.6-27B
A KL-divergence benchmark on bartowski's Qwen3.6-27B GGUF quants (Q8/Q6/Q5) shows KV cache quantization at (q8_0,q8_0) preserves quality almost for free.
CommitBrief — AI code reviews, right in your terminal
A provider-agnostic, local-first CLI that reviews your staged changes, a historic range, or a whole GitHub pull request. Zero telemetry, no server. Free and open source.
commitbrief.comHybrid SWA Efficiency and Improvements with MiMo v2.5
MiMo v2.5 enhances model efficiency with Hybrid SWA, achieving a 60% size reduction.