Agent RL framework for LLM agents: multi-turn reinforcement learning with StarPO and reasoning-collapse diagnostics
-
Updated
Aug 23, 2026 - Python
Agent RL framework for LLM agents: multi-turn reinforcement learning with StarPO and reasoning-collapse diagnostics
RLAnything (ICML 2026) & AutoTool (ICML 2026), DemyAgent: Open-Source RL for LLMs and Agentic Scenarios
Scalable and extensible reinforcement learning for LM agents.
[NeurIPS 2026] FedAgent: A Library for Decentralized Agent Learning. 🏆 Best Paper Award at the AAAI'26 TrustAgent Workshop, 🏆 Outstanding Paper Award in the AAAI'26 PerFM Workshop.
Apprentissage par renforcement pour la régulation des feux de signalisation sur un boulevard typique de Kinshasa
A curated list of training & evaluation environments for LLM/VLM agents (SWE-Gym, GEM, RAGEN, AgentGym, WebArena, OSWorld, ToolBench…). Updated weekly.
An end-to-end framework for Agentic Reinforcement Learning.
Train SLM to use Tools with RL
Project pages for RAGEN, RAGEN-2 (reasoning collapse in agentic RL), and BAGEN (budget-aware LLM agents)
前沿 LLM 中文代码教材:KDA、AttnRes、MSA、mHC、Engram、MTP、Muon、Agent RL、多教师蒸馏、FP4 与本地 Serving · 18 章课程 · 无需云资源
Agent-RL Credit Auditor: CPU-first audit and exact benchmark for credit estimators in agent RL. Explicit estimands, import-isolated Bellman/enumeration oracles, matched budgets, mechanism gates, claim ceilings.
Compiler feedback as process reward for coding agent RL training (Junhao Fu, 2025)
To associate your repository with the agent-rl topic, visit your repo's landing page and select "manage topics."