» Tag
machine-learning
163 postsSparse Policy Selection in RL for LLM Reasoning, Not Capability Learning
Reinforcement learning enhances LLM reasoning by focusing on sparse policy selection rather than teaching new capabilities.
Open Model Enhanced for Scientific Inquiry
The open model developed by Loka and Arcee AI enhances scientific inquiry through tool use and logical reasoning.
NVIDIA Alpamayo 2 Super: 34B Open Vision-Language-Action Model
NVIDIA introduces Alpamayo 2 Super, a 34B open vision-language-action model for robotaxis.
CommitBrief — AI code reviews, right in your terminal
A provider-agnostic, local-first CLI that reviews your staged changes, a historic range, or a whole GitHub pull request. Zero telemetry, no server. Free and open source.
commitbrief.comThinking Machines Launches Inkling Small Open Source AI Model
Thinking Machines introduces Inkling-Small, a compact open-source AI model with impressive performance.
Measuring LLMs’ Ability to Perform Cryptanalysis
A new benchmark evaluates LLMs' cryptanalysis capabilities, revealing new vulnerabilities in cryptographic schemes.
Profiling in PyTorch (Part 3): Attention is All You Profile
Explore profiling the attention mechanism in PyTorch, focusing on performance enhancements through in-place operations.
AI Reverse Engineering Benchmark
AgentRE-Bench assesses AI agents' reverse engineering capabilities. Calibration proves more effective than reasoning depth.
Open-ultra: A Self-Training LLM Routing Proxy
Open-ultra is a self-training LLM routing proxy that reduces costs while maintaining high output quality.
Qwen3.6-35B-A3B on 4× Intel Arc Pro B70: Achieving 200 tok/s
The Qwen3.6-35B-A3B model offers four optimized configurations with 4 Intel Arc Pro B70 GPUs, balancing performance and flexibility.
Code Repair Training Data Created and Evaluation Released
A new code repair dataset enables engineers to test accuracy, improving fix rates from 17% to 40%.