» Tag
scaling-laws
2 postsScaling Laws Explained: From Kaplan to Chinchilla to Overtraining
A breakdown of LLM scaling laws from Kaplan to Chinchilla, and why modern models are deliberately overtrained to cut inference costs.
New study finds language models memorize about 3.6 bits per parameter
Researchers unveil a method to measure LLM memorization capacity, finding GPT-style models store roughly 3.6 bits of information per parameter, with implications for grokking and scaling.
CommitBrief — AI code reviews, right in your terminal
A provider-agnostic, local-first CLI that reviews your staged changes, a historic range, or a whole GitHub pull request. Zero telemetry, no server. Free and open source.
commitbrief.com