» Tag
performance
289 postsWasmi 2.0: Engineering the Fastest WebAssembly Interpreters
Wasmi 2.0 offers a 2.2x speed increase over its predecessor, introducing new features and optimizations crucial for engineers in various applications.
Nvidia and Cerebras: Selling Performance Customers Will Likely Never See
Nvidia and Cerebras showcased their performance metrics at Hot Chips. However, these figures may not reflect real-world usage.
From LLM Inference to Agentic Workloads: Characterization and Implications
Explore how agentic applications are reshaping LLM services and their system behaviors.
CommitBrief — AI code reviews, right in your terminal
A provider-agnostic, local-first CLI that reviews your staged changes, a historic range, or a whole GitHub pull request. Zero telemetry, no server. Free and open source.
commitbrief.comRunning Multiple Linux Kernels Simultaneously on Bare Metal
Mklinux v7.0-mk2 enables running multiple Linux kernels simultaneously on bare metal.
Taffy: A Flexible, High-Performance UI Layout Library
Taffy is a flexible, high-performance UI layout library written in Rust.
Hiding Memory Latency in eBPF: Solutions to Avoid Stalling
Exploring methods to hide memory latency with eBPF and its importance for engineers.
Self-hosting Kimi K3: Cost and Performance Insights
Analysis of cost and performance when self-hosting Kimi K3.
Ninfer: High-performance single-GPU inference
NInfer offers a high-performance C++/CUDA inference engine for RTX 5090.
Frontier-class LLM Inference on a Laptop CPU
cpubrrr surpasses llama.cpp on Apple M4 Max CPU for frontier-class LLMs.
Scaling Live Audio: Insights from QUIC and MoQ Relays
Key insights from scaling Tula Spaces for live audio and lessons for engineers.