Read original ↗
paperarXivTrust 82 · PrimaryPublished 12h agoLive · 1h ago

Don't Drop the BATON: Long-Horizon Robot Manipulation via Agentic Subtask Exploration and Transition-aware Memory

Long-horizon robot manipulation chains many contact-rich skills into one multi-stage task. Vision-language-action (VLA) models increasingly master the individual skills, yet the chain still fails: errors compound beyond the policy's ability to correct, and one subtask silently constrains the next. A promising recipe freezes the VLA and puts an LLM agent in charge: it plans in language, moves in free space with analytic primitives, invokes the VLA only for contact-rich segments, and writes adaptation into language memory. Applied to long horizons, it breaks twice. (1) Competence comes from whol

Lineage graph

Paper → model → repo connections mined from source citations (Tier-1 exact match).

Why these links exist

Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.

  • FuzzySimilar title/name (fuzzy) · 84%Thysrael/Horizon

    Fuzzy title match (0.92): “Don't Drop the BATON: Long-Horizon Robot Manipulation via Ag” ≈ “Thysrael/Horizon”

  • FuzzySimilar title/name (fuzzy) · 59%Fosowl/agenticSeek

    Fuzzy title match (0.73): “Don't Drop the BATON: Long-Horizon Robot Manipulation via Ag” ≈ “Fosowl/agenticSeek”

  • LinkedLinked via arxiv author · 85%Bingxin Xu

    Don't Drop the BATON: Long-Horizon Robot Manipulation via Agentic Subtask Exploration and Transition-aware Memory

  • LinkedLinked via arxiv author · 85%Yuzhang Shang

    Don't Drop the BATON: Long-Horizon Robot Manipulation via Agentic Subtask Exploration and Transition-aware Memory

  • LinkedLinked via arxiv author · 85%Emilio Ferrara

    Don't Drop the BATON: Long-Horizon Robot Manipulation via Agentic Subtask Exploration and Transition-aware Memory

Implements (incoming)

authored (incoming)

Related across the graph

Topics