Agentic Rl
6 items across the graph — tagged with Agentic Rl.
From the graph · 6
🚀 An open-source, hands-on curriculum bridging the gap from basic RL concepts to LLM alignment, RLVR, and advanced Agentic systems.
Agent-R1: Training Powerful LLM Agents with End-to-End Reinforcement Learning
An Asynchronous Reinforcement Learning Engine for Omni-Modal Post-Training at Scale
Curated papers, taxonomy, benchmarks, and decision guides for credit assignment in reasoning and agentic LLM reinforcement learning.
Run a documented subset of verl-style OPD on one consumer GPU—typed config, Parquet prompts, and PEFT scale-out artifacts.
Curated, opinionated index of post-R1 LLM × Reinforcement Learning. Many deep-dive blog posts cross-linked to many papers — GRPO, DAPO, DPO, PPO, RLHF, GSPO, CI…
