Unity-Technologies/ml-agents
The Unity Machine Learning Agents Toolkit (ML-Agents) is an open-source project that enables games and simulations to serve as environments for training intelligent agents using deep reinforcement learning and imitation learning.
Lineage graph
Paper → model → repo connections mined from source citations (Tier-1 exact match).
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.
- FuzzySimilar title/name (fuzzy) · 84%Empowering GUI Agents via Autonomous Experience Exploration and Hindsight Experience Utilization for Task Planning →
“Fuzzy title match (0.92): “Empowering GUI Agents via Autonomous Experience Exploration ” ≈ “Unity-Technologies/ml-agents””
- FuzzySimilar title/name (fuzzy) · 84%MetaSkill-Evolve: Recursive Self-Improvement of LLM Agents via Two-Timescale Meta-Skill Evolution →
“Fuzzy title match (0.92): “MetaSkill-Evolve: Recursive Self-Improvement of LLM Agents v” ≈ “Unity-Technologies/ml-agents””
- FuzzySimilar title/name (fuzzy) · 84%When State Becomes an Attack Surface: State-Semantic Injection in LLM-Driven Embodied Agents →
“Fuzzy title match (0.92): “When State Becomes an Attack Surface: State-Semantic Injecti” ≈ “Unity-Technologies/ml-agents””
- FuzzySimilar title/name (fuzzy) · 84%When Agents Coordinate: Measuring Coordination in Multi-Agent AI Coding →
“Fuzzy title match (0.92): “When Agents Coordinate: Measuring Coordination in Multi-Agen” ≈ “Unity-Technologies/ml-agents””
- FuzzySimilar title/name (fuzzy) · 84%Neurosymbolic Embodied Agents →
“Fuzzy title match (0.92): “Neurosymbolic Embodied Agents” ≈ “Unity-Technologies/ml-agents””
- FuzzySimilar title/name (fuzzy) · 84%StagedWorkspace: A Versioned Workspace for Knowledge-Work Agents →
“Fuzzy title match (0.92): “StagedWorkspace: A Versioned Workspace for Knowledge-Work Ag” ≈ “Unity-Technologies/ml-agents””
- FuzzySimilar title/name (fuzzy) · 84%Policy-Invariant Reward Shaping from LLM Feedback: A Framework for Hybrid RL Agents →
“Fuzzy title match (0.92): “Policy-Invariant Reward Shaping from LLM Feedback: A Framewo” ≈ “Unity-Technologies/ml-agents””
- FuzzySimilar title/name (fuzzy) · 84%On the Fragility of Self-Improving Agents: Variance, Task Order, and Underspecification →
“Fuzzy title match (0.92): “On the Fragility of Self-Improving Agents: Variance, Task Or” ≈ “Unity-Technologies/ml-agents””
