repoGitHubTrust 82 · PrimaryPublished 1mo agoLive · 24d ago
EvolvingLMMs-Lab/LLaVA-OneVision-2
Fully Open Framework for Democratized Multimodal Training
Lineage graph
Paper → model → repo connections mined from source citations (Tier-1 exact match).
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.
- PossiblePossibly related (embedding) · 64%Ask, Solve, Generate: Self-Evolving Unified Multimodal Understanding and Generation via Self-Consistency Rewards →
- PossiblePossibly related (embedding) · 58%Anyone looking into the new MARS2 Workshop/Competition @ ECCV 2026? I saw Tec-do posting it. [D] →
- PossiblePossibly related (embedding) · 57%Paying More Attention to Visual Tokens in Self-Evolving Large Multimodal Models →
- PossiblePossibly related (embedding) · 55%Orca: The World is in Your Mind →
- PossiblePossibly related (embedding) · 53%ManimAgent: Self-Evolving Multimodal Agents for Visual Education →
- PossiblePossibly related (embedding) · 46%SAGEAgent: A Self-Evolving Agent for Cost-Aware Modality Acquisition in Multimodal Survival Prediction →
- PossiblePossibly related (embedding) · 47%ALICE: Learning a General-Purpose Pathology Foundation Model from Vision, Vision-Language, and Slide-Level Experts →
- PossiblePossibly related (embedding) · 53%LoRA-Based Cascaded Multimodal Fusion for Action Recognition in Medical Training Environments →
Implements
paperAsk, Solve, Generate: Self-Evolving Unified Multimodal Understanding and Generation via Self-Consistency RewardspaperPaying More Attention to Visual Tokens in Self-Evolving Large Multimodal ModelspaperOrca: The World is in Your MindpaperManimAgent: Self-Evolving Multimodal Agents for Visual Education
Covers
Implements (incoming)
paperSAGEAgent: A Self-Evolving Agent for Cost-Aware Modality Acquisition in Multimodal Survival PredictionpaperALICE: Learning a General-Purpose Pathology Foundation Model from Vision, Vision-Language, and Slide-Level ExpertspaperLoRA-Based Cascaded Multimodal Fusion for Action Recognition in Medical Training EnvironmentspaperHy-Embodied-VLM-1.0: Efficient Physical-World Agents
Covers (incoming)
Related across the graph
paperSAGEAgent: A Self-Evolving Agent for Cost-Aware Modality Acquisition in Multimodal Survival PredictionpaperALICE: Learning a General-Purpose Pathology Foundation Model from Vision, Vision-Language, and Slide-Level ExpertspaperHy-Embodied-VLM-1.0: Efficient Physical-World AgentspaperManimAgent: Self-Evolving Multimodal Agents for Visual EducationpaperLoRA-Based Cascaded Multimodal Fusion for Action Recognition in Medical Training EnvironmentsnewsFLUX 3 - Real World Models: Towards Multimodal Flow Models as the Backbone of Visual IntelligencenewsAnyone looking into the new MARS2 Workshop/Competition @ ECCV 2026? I saw Tec-do posting it. [D]paperPaying More Attention to Visual Tokens in Self-Evolving Large Multimodal ModelspaperAsk, Solve, Generate: Self-Evolving Unified Multimodal Understanding and Generation via Self-Consistency RewardspaperOrca: The World is in Your Mind
