» Tag
mixture-of-experts
4 postsTurboFieldfare runs Gemma 4 26B MoE model in 2GB RAM on any Mac
TurboFieldfare is an open-source Swift/Metal runtime that streams MoE experts to run Gemma 4 26B in just 2GB RAM on 8GB Apple Silicon Macs.
Moonshot's 2.8T-Parameter Kimi K3 Runs on a GPU-less Mini PC
A 2.8T-parameter open-weight Kimi K3 model ran on a GPU-less mini PC using disk-streamed MoE experts and native MXFP4 quantization.
Moonshot AI Ships Kimi K3: A 2.8T-Parameter Open MoE Model
Moonshot AI's Kimi K3 is a 2.8T-parameter open MoE model with 1M context and new attention and routing architectures for long-context inference.
CommitBrief — AI code reviews, right in your terminal
A provider-agnostic, local-first CLI that reviews your staged changes, a historic range, or a whole GitHub pull request. Zero telemetry, no server. Free and open source.
commitbrief.comLingBot-Video: A New Open-Source MoE Model for Embodied Video Generation
Robbyant's open-source MoE video model LingBot-Video tops the RBench leaderboard, shipping Apache 2.0 code, weights, and inference tooling.