Read original ↗
paperarXivTrust 82 · PrimaryPublished 15d agoLive · 13d ago

SAM-MT: Real-Time Interactive Multi-Target Video Segmentation

Modern Video Object Segmentation (VOS) involves tracking and segmenting user-specified targets. While recent approaches have achieved remarkable performance in single-target scenarios, extending them to multi-target settings typically involves replicating the single-target processing for each individual object, resulting in reduced frame rates (FPS) with unbounded latency as target count increases. Built upon Segment Anything 2 (SAM2), we propose SAM-MT, which addresses this by transforming the model into an interactive framework for real-time Multi-Target video segmentation. SAM-MT uses expli

Lineage graph

Paper → model → repo connections mined from source citations (Tier-1 exact match).

Why these links exist

Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.

  • PossiblePossibly related (embedding) · 50%Somnusochi/VLM-AutoYOLO
  • LinkedLinked via arxiv author · 85%Ruiqi Shen

    SAM-MT: Real-Time Interactive Multi-Target Video Segmentation

  • LinkedLinked via arxiv author · 85%Chang Liu

    SAM-MT: Real-Time Interactive Multi-Target Video Segmentation

  • LinkedLinked via arxiv author · 85%Henghui Ding

    SAM-MT: Real-Time Interactive Multi-Target Video Segmentation

Implements

authored (incoming)

Related across the graph

Topics