» Tag
hardware
34 postsUnified Memory: Why Mini PCs Run 70B Models a Big GPU Can't
Unified-memory mini PCs like AMD's Strix Halo can load 70B-parameter models that a $2,000 RTX 5090 cannot fit. Here's why capacity and bandwidth pull in opposite directions.
OpenAI and Broadcom Unveil Jalapeño, an LLM Inference Chip
OpenAI and Broadcom unveiled Jalapeño, a custom LLM inference chip built in nine months, set for gigawatt-scale data center deployment starting in 2026.
Self-hosting Kimi K3: Cost and Performance Insights
Analysis of cost and performance when self-hosting Kimi K3.
CommitBrief — AI code reviews, right in your terminal
A provider-agnostic, local-first CLI that reviews your staged changes, a historic range, or a whole GitHub pull request. Zero telemetry, no server. Free and open source.
commitbrief.comHow I Fixed Power Management on a Mainline OnePlus 3T
Insights into the power management issues and solutions for the OnePlus 3T.
Testing LLM Concurrency on Consumer Hardware (RTX 5060)
LLM concurrency tests on RTX 5060 yield crucial insights for engineers.
Understanding NPU TOPS and Memory Bandwidth for Local LLM Performance
NPU TOPS and memory bandwidth are critical factors affecting local LLM performance.
Maurice Wilkes: Transforming Hardware Problems into Software Solutions
Maurice Wilkes revolutionized computing with microprogramming, enabling software changes without hardware modifications, enhancing reusability.
Running Gemma 4 26B on a 13-Year-Old Xeon Without a GPU
Explore how to run Gemma 4 26B on an old Xeon server without a GPU.
Benchmarking a Personal AI PC on Battery, Thermals, and Sleep Recovery
Evaluate battery, thermal, and sleep recovery performance of NVIDIA's RTX Spark for personal AI PCs.
NVIDIA Ising: The AI Models Tackling Quantum Computing's Scaling Barrier
NVIDIA unveils Ising, open-source AI models for quantum processor calibration and error correction, integrated with CUDA-Q to unlock large-scale quantum computing.