» Tag
language-models
14 postsSparse Policy Selection in RL for LLM Reasoning, Not Capability Learning
Reinforcement learning enhances LLM reasoning by focusing on sparse policy selection rather than teaching new capabilities.
Trie-Based Memory Efficiency for Long-Context Compression
SALT compresses long documents for language models, reducing memory and compute time.
A Cliché-Resistant Gemma Model for Storytelling
Gemma-4-26B-A4B-StyleTune-V2 is a language model designed for cliché-free storytelling.
CommitBrief — AI code reviews, right in your terminal
A provider-agnostic, local-first CLI that reviews your staged changes, a historic range, or a whole GitHub pull request. Zero telemetry, no server. Free and open source.
commitbrief.comWhy Long Prompts Diminish AI Performance and How to Improve Them
Discover how long prompts affect AI outputs and learn effective compression techniques.
German AI Consortium Unveils Soofi S: An Open 30B Model
Soofi S, developed by the German AI consortium, achieves top benchmark scores.
Empero AI Launches Qwythos-9B-v2: Looping Issues Resolved and Robustness Enhanced
Empero AI's Qwythos-9B-v2 resolves looping issues and enhances robustness, offering a 1M-token context for advanced applications.
MET: Advancing Multilingual Moral Reasoning with Culture-Aware Theory
MET introduces innovative frameworks for enhancing multilingual moral reasoning in language models.
Why Your Prompts Fail and How to Fix Them
Discover why your prompts fail and how to enhance their effectiveness.
GLM-5.3 arrives with advanced cyber capabilities and finds vulnerability in Cursor
Z.ai's GLM-5.3 model has launched, finding a serious vulnerability in Cursor and showcasing advanced cybersecurity capabilities.
No cloud, no GPUs: Liquid AI's LFM2.5-2.6B model empowers edge devices
Liquid AI introduces LFM2.5-2.6B, a model that runs on local hardware without cloud or GPU dependencies.