» Tag
gpu
91 postsQwen 3.6 27B Model on 16GB VRAM Achieves 20 t/s Performance
Insights on the Qwen 3.6 27B model configuration and performance testing on 16GB VRAM.
Benchmarking 15 'E-Waste' GPUs with Modern Workloads
An evaluation of decommissioned NVIDIA Tesla GPUs' performance with modern workloads. Offers crucial insights for low-cost GPU node construction.
Dgxtop: Rust-Based Monitor for NVIDIA DGX Systems
Dgxtop offers real-time GPU and system monitoring for NVIDIA DGX systems.
CommitBrief — AI code reviews, right in your terminal
A provider-agnostic, local-first CLI that reviews your staged changes, a historic range, or a whole GitHub pull request. Zero telemetry, no server. Free and open source.
commitbrief.comRunning Multiple ComfyUI Instances on a Single GPU
Explore ways to run multiple instances of ComfyUI on a single GPU.
Unified Memory: Why Mini PCs Run 70B Models a Big GPU Can't
Unified-memory mini PCs like AMD's Strix Halo can load 70B-parameter models that a $2,000 RTX 5090 cannot fit. Here's why capacity and bandwidth pull in opposite directions.
The First Open-Source Agent Skills Collection for AMD ROCm
While NVIDIA has 428+ agent skills on skills.sh, AMD had none. This new open-source project delivers 10 production-ready skills for ROCm GPU workflows.
Flux: Compile an LLM to Your Hardware and Serve It
Flux creates the optimal plan for LLM inference on your hardware.
Training a Language Model End-to-End in Rust: An Experience Report
An experience report on training a language model in Rust and the encountered issues.
The AI Data-Centre Bust Will Look Like a Boom
The growing demand for AI boosts the importance of custom chips, but GPU purchases with debt could trigger a bust in data centers.
Making Long-Running NVIDIA TensorRT Engine Builds Observable and Cancelable
Learn to use the IProgressMonitor API for observable and cancelable TensorRT engine builds.