AgentGym-RL
paper Your tags
Your notes
Framework for training LLM agents for long-horizon multi-turn decision-making via reinforcement learning across diverse environments, introducing the ScalingInter-RL recipe that progressively scales agent-environment interaction during training. From Fudan NLP (Tao Gui, Zhiheng Xi) with ByteDance.