Agent RL framework for LLM agents: multi-turn reinforcement learning with StarPO and reasoning-collapse diagnostics
-
Updated
Aug 23, 2026 - Python
Agent RL framework for LLM agents: multi-turn reinforcement learning with StarPO and reasoning-collapse diagnostics
RLAnything (ICML 2026) & AutoTool (ICML 2026), DemyAgent: Open-Source RL for LLMs and Agentic Scenarios
Scalable and extensible reinforcement learning for LM agents.
Apprentissage par renforcement pour la régulation des feux de signalisation sur un boulevard typique de Kinshasa
A curated list of training & evaluation environments for LLM/VLM agents (SWE-Gym, GEM, RAGEN, AgentGym, WebArena, OSWorld, ToolBench…). Updated weekly.
Train SLM to use Tools with RL
Project pages for RAGEN, RAGEN-2 (reasoning collapse in agentic RL), and BAGEN (budget-aware LLM agents)
GRPO integration for multi-turn retail agents with Tau2 environments, VERL AgentLoop, tool rollouts, and terminal rewards.
Agent-RL Credit Auditor: CPU-first audit and exact benchmark for credit estimators in agent RL. Explicit estimands, import-isolated Bellman/enumeration oracles, matched budgets, mechanism gates, claim ceilings.
Compiler feedback as process reward for coding agent RL training (Junhao Fu, 2025)
To associate your repository with the agent-rl topic, visit your repo's landing page and select "manage topics."