RAGEN: train and evaluate LLM agents using multi-turn RL

03/05/2025 11 min

Listen "RAGEN: train and evaluate LLM agents using multi-turn RL"

Descargar episodio Ver en sitio original

Episode Synopsis

RAGEN is a modular system for training and evaluating LLM agents using multi-turn reinforcement learning. Built on the StarPO framework, it implements the full training loop including rollout generation, reward assignment, and trajectory optimization. RAGEN serves as research infrastructure to analyze LLM agent training dynamics, focusing on challenges like stability, generalization, and the emergence of reasoning in interactive environments.

More episodes of the podcast Large Language Model (LLM) Talk

Kimi K2 22/07/2025

Mixture-of-Recursions (MoR) 18/07/2025

MeanFlow 10/07/2025

Mamba 10/07/2025

LLM Alignment 14/06/2025

Why We Think 20/05/2025

Deep Research 12/05/2025

vLLM 04/05/2025

Qwen3: Thinking Deeper, Acting Faster 04/05/2025

DeepSeek-Prover-V2 01/05/2025

Ver todos los episodios

ZARZA We are Zarza, the prestigious firm behind major projects in information technology.

RAGEN: train and evaluate LLM agents using multi-turn RL

Listen "RAGEN: train and evaluate LLM agents using multi-turn RL"

Episode Synopsis

More episodes of the podcast Large Language Model (LLM) Talk

Free Internet, a prediction in Nostradamus style

Personnel recruitment via Web

Bandwidth: Broadband or Narrowband?

Personnel recruitment via Web

Deep web or Invisible Internet

Subdomains, a glance with the experts!

Free Internet, a prediction in Nostradamus style

Educational Technology: From traditional to digital

Localhost, there’s no place like 127.0.0.1

Googling with breathtaking tricks you ignore

Gray Hat Hacking, those with ambiguous ethics…

Internet Predators on the prowl

Dot COM: The Internet’s dominant TLD