Skip to content
Artwork for Large Language Model (LLM) Talk
Large Language Model (LLM) Talk · May 3, 2025 · 11 min

RAGEN: train and evaluate LLM agents using multi-turn RL

RAGEN is a modular system for training and evaluating LLM agents using multi-turn reinforcement learning. Built on the StarPO framework, it implements the full training loop including rollout generation, reward assignment, and trajectory optimization. RAGEN serves as research infrastructure to analyze LLM agent training dynamics, focusing on challenges like stability, generalization, and the emergence of reasoning in interactive environments.

0:00-11:56

transcript

No transcript — this publisher did not publish one.

show notes

RAGEN is a modular system for training and evaluating LLM agents using multi-turn reinforcement learning. Built on the StarPO framework, it implements the full training loop including rollout generation, reward assignment, and trajectory optimization. RAGEN serves as research infrastructure to analyze LLM agent training dynamics, focusing on challenges like stability, generalization, and the emergence of reasoning in interactive environments.

more episodes

All episodes