{"repo":"langfengQ/verl-agent","free":true,"listed":false,"github":"https://github.com/langfengQ/verl-agent","clone":"git clone https://github.com/langfengQ/verl-agent.git","description":"verl-agent is an extension of veRL, designed for training LLM/VLM agents via RL. verl-agent is also the official code for paper \"Group-in-Group Policy Optimization for LLM Agent Training\"","language":"Python","stars":2228,"topics":["llm-agents","llm-training","reinforcement-learning","large-language-models","deepseek-r1","grpo","agent-framework","gigpo"],"license":"Apache-2.0","category":"ai-agents","readme_excerpt":"Group-in-Group Policy Optimization for LLM Agent Training NeurIPS 2025 &nbsp; &nbsp; &nbsp; &nbsp; &nbsp; &nbsp; &nbsp; verl-agent is an extension of veRL, specifically designed for training large language model (LLM) agents via reinforcement learning (RL) . Unlike prior approaches that simply concatenate full interaction histories, verl-agent proposes step-independent multi-turn rollout mechanism , which allows for fully customizable per-step input structures, history management, and memory modules. This design makes verl-agent highly scalable for very long-horizon, multi-turn RL training (e.g., tasks in ALFWorld can require up to 50 steps to complete). verl-agent provides a diverse set of RL algorithms (including our new algorithm GiGPO) and a rich suite of agent environments , enabling the development of reasoning agents in both visual and text-based tasks. News - [2026.05] GraphGPO accepted at ICML 2026! 🎉🎉🎉 [Paper] [Code] - [2026.02] HGPO accepted at ICLR 2026! 🎉🎉🎉 [Paper] [Code] - [2026.02] 🔥 We open-source Dr. MAS, which supports stable end-to-end RL post-training of multi-agent LLM systems ! [Paper] [Code] - [2025.12] Qwen3-VL is supported! See example here. - [2025.09] GiGPO is now supported by ROLL! [Document] [Train Curves]. - [2025.09] verl-agent -style training pipeline is now supported by OpenManus-RL! - [2025.09] GiGPO accepted at NeurIPS 2025! 🎉🎉🎉 - [2025.08] Add Search-R1 experiments and similarity-based GiGPO ! Check out GiGPO's superior performanc","default_branch":null,"files":null,"tree":[],"storefront":"/r/langfengQ","claimed":false,"request_supported":{"post":"https://gitbuyer.com/r/langfengQ/verl-agent/request-supported","requests":0},"note":"indexed from public GitHub; nothing is for sale on this page. Clone it from GitHub. Paid listings live at /search."}