repoGitHubTrust 82 · PrimaryPublished 1mo agoLive · 24d ago
redai-infra/Relax
An Asynchronous Reinforcement Learning Engine for Omni-Modal Post-Training at Scale
Lineage graph
Paper → model → repo connections mined from source citations (Tier-1 exact match).
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.
- PossiblePossibly related (embedding) · 53%Joint Learning of Experiential Rules and Policies for Large Language Model Agents →
- PossiblePossibly related (embedding) · 51%Z-1: Efficient Reinforcement Learning for Vision-Language-Action Models →
- PossiblePossibly related (embedding) · 50%RL without TD learning →
- PossiblePossibly related (embedding) · 50%Is One Layer Enough? Training A Single Transformer Layer Can Match Full-Parameter RL Training →
- PossiblePossibly related (embedding) · 49%Ask, Solve, Generate: Self-Evolving Unified Multimodal Understanding and Generation via Self-Consistency Rewards →
- PossiblePossibly related (embedding) · 46%CompactionRL: Reinforcement Learning with Context Compaction for Long-Horizon Agents →
- PossiblePossibly related (embedding) · 46%Multi-Modal, Multi-Environment Machine Teaching for Robust Reward Learning →
- PossiblePossibly related (embedding) · 52%Mach-Mind-4-Flash Technical Report →
Implements
paperJoint Learning of Experiential Rules and Policies for Large Language Model AgentspaperZ-1: Efficient Reinforcement Learning for Vision-Language-Action ModelspaperIs One Layer Enough? Training A Single Transformer Layer Can Match Full-Parameter RL TrainingpaperAsk, Solve, Generate: Self-Evolving Unified Multimodal Understanding and Generation via Self-Consistency Rewards
Covers
Implements (incoming)
Covers (incoming)
Related across the graph
newsRL without TD learningpaperActive Offline-to-Online Reinforcement LearningpaperMulti-Modal, Multi-Environment Machine Teaching for Robust Reward LearningpaperMach-Mind-4-Flash Technical ReportnewsSeeking collaborators for scaling and independent evaluation of a new recurrent language model architecture (preprint + code) [R]paperJoint Learning of Experiential Rules and Policies for Large Language Model AgentspaperCompactionRL: Reinforcement Learning with Context Compaction for Long-Horizon AgentspaperIs One Layer Enough? Training A Single Transformer Layer Can Match Full-Parameter RL TrainingpaperZ-1: Efficient Reinforcement Learning for Vision-Language-Action ModelspaperAsk, Solve, Generate: Self-Evolving Unified Multimodal Understanding and Generation via Self-Consistency Rewards
