SAM-MT: Real-Time Interactive Multi-Target Video Segmentation
Modern Video Object Segmentation (VOS) involves tracking and segmenting user-specified targets. While recent approaches have achieved remarkable performance in single-target scenarios, extending them to multi-target settings typically involves replicating the single-target processing for each individual object, resulting in reduced frame rates (FPS) with unbounded latency as target count increases. Built upon Segment Anything 2 (SAM2), we propose SAM-MT, which addresses this by transforming the model into an interactive framework for real-time Multi-Target video segmentation. SAM-MT uses expli
Lineage graph
Paper → model → repo connections mined from source citations (Tier-1 exact match).
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.
- PossiblePossibly related (embedding) · 50%Somnusochi/VLM-AutoYOLO →
- LinkedLinked via arxiv author · 85%Ruiqi Shen →
“SAM-MT: Real-Time Interactive Multi-Target Video Segmentation”
- LinkedLinked via arxiv author · 85%Chang Liu →
“SAM-MT: Real-Time Interactive Multi-Target Video Segmentation”
- LinkedLinked via arxiv author · 85%Henghui Ding →
“SAM-MT: Real-Time Interactive Multi-Target Video Segmentation”
