» Tag
reinforcement-learning
14 postsRL-Training Agent Developed for Model Training
Learn about the new pipeline developed for model training using an AI agent.
GEPA: Reflective Prompt Evolution Can Outperform Reinforcement Learning
GEPA enhances LLM learning through language reflection, outperforming reinforcement learning methods.
The Comprehensive Guide to Agentic AI: An End-to-End Practitioner Resource
A new arXiv book covers the agentic AI stack under one roof, from LLM fundamentals to multi-agent architectures. A practical reference for engineers.
CommitBrief — AI code reviews, right in your terminal
A provider-agnostic, local-first CLI that reviews your staged changes, a historic range, or a whole GitHub pull request. Zero telemetry, no server. Free and open source.
commitbrief.comFinetuning a Reasoning LLM with Supervised or Reinforcement Learning?
Critical insights on training data representation and loss management in LLM finetuning.