» Tag
energy-efficiency
2 postsPersistent State Machines Recast LLM Attention as Hardware FSMs
A PSM framework recasts LLM attention as deterministic hardware state machines, synthesized in Vivado with an estimated 0.0267 pJ/op dynamic energy.
Benchmarking Inference Energy Costs of LLM: A LLaMA Study
Researchers benchmark LLaMA model sizes on V100 and A100 GPUs to analyze the energy and compute costs of LLM inference at scale.
CommitBrief — AI code reviews, right in your terminal
A provider-agnostic, local-first CLI that reviews your staged changes, a historic range, or a whole GitHub pull request. Zero telemetry, no server. Free and open source.
commitbrief.com