repoGitHubTrust 82 · PrimaryPublished 13d agoLive · 4d ago
lubludrova/rl-handbook
A comprehensive guide to Reinforcement Learning
Lineage graph
Paper → model → repo connections mined from source citations (Tier-1 exact match).
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.
- PossiblePossibly related (embedding) · 55%RLHF →
- PossiblePossibly related (embedding) · 52%[2607.07508] Single-Rollout Asynchronous Optimization for Agentic Reinforcement Learning →
- PossiblePossibly related (embedding) · 52%Maxims for machines: operationalizing Kant’s universal law via inverse reinforcement learning - Springer Nature Link →
- PossiblePossibly related (embedding) · 52%The Little Book of Reinforcement Learning →
- PossiblePossibly related (embedding) · 51%Best practices for multi-turn reinforcement learning in Amazon SageMaker AI →
Related to
Covers
news[2607.07508] Single-Rollout Asynchronous Optimization for Agentic Reinforcement LearningnewsMaxims for machines: operationalizing Kant’s universal law via inverse reinforcement learning - Springer Nature LinknewsThe Little Book of Reinforcement LearningnewsBest practices for multi-turn reinforcement learning in Amazon SageMaker AI
Related across the graph
news[2607.07508] Single-Rollout Asynchronous Optimization for Agentic Reinforcement LearningnewsMaxims for machines: operationalizing Kant’s universal law via inverse reinforcement learning - Springer Nature LinknewsBest practices for multi-turn reinforcement learning in Amazon SageMaker AIglossary_termRLHFnewsThe Little Book of Reinforcement Learning
