repoGitHubTrust 82 · PrimaryPublished 25d agoLive · 24d ago
inclusionAI/AReno
An easy-to-use, fast toolkit to scale up RL post-training on a single node.
Lineage graph
Paper → model → repo connections mined from source citations (Tier-1 exact match).
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.
- PossiblePossibly related (embedding) · 54%Seeking collaborators for scaling and independent evaluation of a new recurrent language model architecture (preprint + code) [R] →
- PossiblePossibly related (embedding) · 51%RL without TD learning →
- PossiblePossibly related (embedding) · 47%Reproducing OpenAI’s “persistently beneficial models” - GRPO trait install barely moves. Ideas? [P] [R] →
