DLR-RM/stable-baselines3
PyTorch version of Stable Baselines, reliable implementations of reinforcement learning algorithms.
Lineage graph
Paper → model → repo connections mined from source citations (Tier-1 exact match).
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.
- FuzzySimilar title/name (fuzzy) · 59%MM-Spectrum: Multimodal Multi-spectral Molecular Structural Elucidation with a Stable MoE Framework →
“Fuzzy title match (0.73): “MM-Spectrum: Multimodal Multi-spectral Molecular Structural ” ≈ “DLR-RM/stable-baselines3””
- FuzzySimilar title/name (fuzzy) · 59%UE5M3 FP4 Block Scaling for Stable Language Model Pretraining →
“Fuzzy title match (0.73): “UE5M3 FP4 Block Scaling for Stable Language Model Pretrainin” ≈ “DLR-RM/stable-baselines3””
- FuzzySimilar title/name (fuzzy) · 59%Stable and Scalable Bundle Adjustment of Holistic 3D Structures →
“Fuzzy title match (0.73): “Stable and Scalable Bundle Adjustment of Holistic 3D Structu” ≈ “DLR-RM/stable-baselines3””
- PossiblePossibly related (embedding) · 57%[2607.07508] Single-Rollout Asynchronous Optimization for Agentic Reinforcement Learning →
- PossiblePossibly related (embedding) · 56%The Little Book of Reinforcement Learning →
- FuzzySimilar title/name (fuzzy) · 59%S3: Stable Subgoal Selection by Constraining Uncertainty of Coarse Dynamics in Hierarchical Reinforcement Learning →
“Fuzzy title match (0.73): “S3: Stable Subgoal Selection by Constraining Uncertainty of ” ≈ “DLR-RM/stable-baselines3””
- PossiblePossibly related (embedding) · 49%The HydroGym reinforcement learning platform for fluid dynamics - nature.com →
- PossiblePossibly related (embedding) · 46%Aftab Improves Reinforcement Learning With 86% Greater Probability Of Success - Quantum Zeitgeist →
