Agent RL framework for LLM agents: multi-turn reinforcement learning with StarPO and reasoning-collapse diagnostics
-
Updated
Aug 23, 2026 - Python
Agent RL framework for LLM agents: multi-turn reinforcement learning with StarPO and reasoning-collapse diagnostics
RLAnything (ICML 2026) & AutoTool (ICML 2026), DemyAgent: Open-Source RL for LLMs and Agentic Scenarios
Scalable and extensible reinforcement learning for LM agents.
Apprentissage par renforcement pour la régulation des feux de signalisation sur un boulevard typique de Kinshasa
A curated list of training & evaluation environments for LLM/VLM agents (SWE-Gym, GEM, RAGEN, AgentGym, WebArena, OSWorld, ToolBench…). Updated weekly.
Train SLM to use Tools with RL
Project pages for RAGEN, RAGEN-2 (reasoning collapse in agentic RL), and BAGEN (budget-aware LLM agents)
Agent-RL Credit Auditor: CPU-first audit and exact benchmark for credit estimators in agent RL. Explicit estimands, import-isolated Bellman/enumeration oracles, matched budgets, mechanism gates, claim ceilings.
Compiler feedback as process reward for coding agent RL training (Junhao Fu, 2025)
Add a description, image, and links to the agent-rl topic page so that developers can more easily learn about it.
To associate your repository with the agent-rl topic, visit your repo's landing page and select "manage topics."