Reward-Guided Autoregressive Graph Generation for Efficient Multi-Agent Communication Topology Design
LLM-based Multi-Agent Systems (MAS) achieve strong performance on complex reasoning tasks by coordinating multiple agents, but at the cost of substantial token consumption. Recent work on automatic topology design, ARG-Designer, has reframed this problem as autoregressive graph generation. However, its training objective provides no explicit incentive for the model to generate sparse and efficient topologies. We address this limitation by introducing a Reward-Guided Autoregressive Graph Generation (RGA-Designer) inspired by Reinforcement Learning from Human Feedback (RLHF). We train a reward m
Lineage graph
Paper → model → repo connections mined from source citations (Tier-1 exact match).
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.
- FuzzySimilar title/name (fuzzy) · 59%AgentCore-8B →
“Fuzzy title match (0.73): “Reward-Guided Autoregressive Graph Generation for Efficient ” ≈ “AgentCore-8B””
- PossiblePossibly related (embedding) · 52%Optimising LMAPF guidance graphs using Evolutionary algorithms: Advice needed [R] →
- FuzzySimilar title/name (fuzzy) · 87%SWE-agent/SWE-agent →
“Fuzzy title match (0.94): “Reward-Guided Autoregressive Graph Generation for Efficient ” ≈ “SWE-agent/SWE-agent””
- FuzzySimilar title/name (fuzzy) · 87%zhayujie/CowAgent →
“Fuzzy title match (0.94): “Reward-Guided Autoregressive Graph Generation for Efficient ” ≈ “zhayujie/CowAgent””
- FuzzySimilar title/name (fuzzy) · 66%open-multi-agent/open-multi-agent →
“Fuzzy title match (0.78): “Reward-Guided Autoregressive Graph Generation for Efficient ” ≈ “open-multi-agent/open-multi-agent””
- FuzzySimilar title/name (fuzzy) · 59%tirth8205/code-review-graph →
“Fuzzy title match (0.73): “Reward-Guided Autoregressive Graph Generation for Efficient ” ≈ “tirth8205/code-review-graph””
- FuzzySimilar title/name (fuzzy) · 59%NousResearch/hermes-agent →
“Fuzzy title match (0.73): “Reward-Guided Autoregressive Graph Generation for Efficient ” ≈ “NousResearch/hermes-agent””
- LinkedLinked via arxiv author · 85%Poomphob Suwannapichat →
“Reward-Guided Autoregressive Graph Generation for Efficient Multi-Agent Communication Topology Design”
