repoGitHubTrust 82 · PrimaryPublished 1mo agoLive · 24d ago
hiyouga/LlamaFactory
Unified Efficient Fine-Tuning of 100+ LLMs & VLMs (ACL 2024)
Lineage graph
Paper → model → repo connections mined from source citations (Tier-1 exact match).
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.
- PossiblePossibly related (embedding) · 53%Fine-tune a small model on your own data →
- PossiblePossibly related (embedding) · 53%Beyond LoRA: Can you beat the most popular fine-tuning technique? →
- PossiblePossibly related (embedding) · 50%OpenAI and Broadcom announce chip designed for LLM inference at scale →
- PossiblePossibly related (embedding) · 48%Evaluate a model properly →
- PossiblePossibly related (embedding) · 47%OpenAI and Broadcom unveil LLM-optimized inference chip →
- FuzzyOverlapping authors or contributors · 62%From RGB Generation to Dense Field Readout: Pixel-Space Dense Prediction with Text-to-Image Models →
“Shared author/contributor keys: lin”
- FuzzyOverlapping authors or contributors · 62%Towards Enhancing 3D Spatial Reasoning in Medical Multimodal Large Language Models →
“Shared author/contributor keys: lin”
- FuzzyOverlapping authors or contributors · 62%MonoIR-RS: Infrared Remote Sensing Vision-Language Learning with CLIP and VLM Adaptation →
“Shared author/contributor keys: lin”
Related to
Covers
Implements
paperFrom RGB Generation to Dense Field Readout: Pixel-Space Dense Prediction with Text-to-Image ModelspaperTowards Enhancing 3D Spatial Reasoning in Medical Multimodal Large Language ModelspaperMonoIR-RS: Infrared Remote Sensing Vision-Language Learning with CLIP and VLM AdaptationpaperPost-Training Shifts Confidence: A Three-Stage Analysis of How SFT, RL, and OPD Shape Pre-, Intra-, and Post-CoT CalibrationpaperPanoWorld: Real-World Panoramic GenerationpaperDynaKRAG: A Unified Framework for Learnable Evidence Control in Multi-Hop Retrieval-Augmented GenerationpaperGatedLinear: Adaptive Routing of Complementary Linear Bases for Time Series ForecastingpaperWat3R: Underwater 3D Geometry Learning without AnnotationspaperLongVQUBench: Benchmarking Long-Term Video Quality Understanding of Vision-Language ModelspaperSynthetic-to-Real Translation for Class-Agnostic Motion PredictionpaperDominoTree: Conditional Tree-Structured Drafting with Domino for Speculative DecodingpaperStable Self-Modulating Quantum Fast-Weight Programmers with Bounded Memory GatespaperReCal3R: Reliability-Calibrated Learning Rates for Streaming 3D ReconstructionpaperLLM Agents for Deliberative Collaboration: A Study on Joint Decision Making Under Partial ObservabilitypaperSMetric: Rethink LLM Scheduling for Serving Agents with Balanced Session-centric SchedulingpaperCan LLMs Write Reliable Rubrics? A Meta-Evaluation for Experiment ReproductionpaperSIVA-RL: Sensitivity-Invariance Visual Alignment for Multimodal Reinforcement LearningpaperMM-IssueLoc: A Controlled Benchmark for Evaluating Visual Evidence in Multimodal Repository-Level Issue LocalizationpaperScaling Behavior Foundation Model for Humanoid RobotspaperBeyond Single Expert: Harmonizing Diverse Visual Priors in MLLMs for Spatial UnderstandingpaperWeakly-Supervised RGB-D Salient Object Detection via SAM-driven Pseudo Annotation and State Space Interaction-based DiffusionpaperOn Success and Simplicity: A Second Look at Transferable Vision-Language Attack PipelinepaperChartGenEval: Corruption-Tested Multi-Dimensional Feedback for Rhythm-Game Chart GenerationpaperEAGLE-360: Embodied Active Global-to-Local Exploration in 360$^\circ$paperRethinking Quantum Continual Learning with Quantum Fisher InformationpaperDepthART: Scaling Foundation Monocular Depth to Tiny ModelspaperCross-Coordinate Correspondence Pruning for Image-to-Point Cloud RegistrationpaperAutoregressive B-Rep Shape Generation with Parametric SurfacespaperSparse Evidence Can Suffice: Agentic Evidence Seeking for Multimodal Video Misinformation DetectionpaperPPL-Factory: Task-Aware and Budget-Aware Data Selection from Language Modeling to ReasoningpaperVGOcc: Learning Visual-Geometric Gaussians for Vision-Centric 3D Driving Occupancy PredictionpaperWhat Transfers Under Source Shift? Definitions, Examples, and Fine-Tuning for Climate Disclosure ClassificationpaperMulti-Modal, Multi-Environment Machine Teaching for Robust Reward LearningpaperMemOps: Benchmarking Lifecycle Memory Operations in Long-Horizon ConversationspaperCR-Refiner: An Object-Centric Optimal Transport Reranker for Edit-Conditioned 3D Scene RetrievalpaperGATE-3D: Geometry-Aware Test-time Adaptive Reranking for Open-Set 3D Shape RetrievalpaperWavefront Parallelization for Efficient Learned Image CompressionpaperFlash EQ-Linear: Accelerating Equivariant Linear Layers via Group-wise Discrete Fourier TransformpaperTexture++: Elevating 3D Asset Texture Resolution with a Region-Aware Diffusion ModelpaperMedGame: Storytelling Gamification Empowered by Large Language Models for Medical EducationpaperLocalize, Then Reason: Visual Latent Structural Reasoning for Molecular Properties and EditspaperHumanTracker: Towards Comprehensive and Human-Aligned Motion Tracking BenchmarkpaperSNM-VFI: Symmetric Nonlinear Motion-Guided Generative Video Frame InterpolationpaperFrom Style Replication to Style Exploration: Enabling Art Style Exploration with Analyze-Experiment-Resituate FrameworkpaperCPI-Bench: A Comprehensive,Practical and Intelligent Benchmark for Real-World Image EditingpaperScalable Visual Pretraining for Language IntelligencepaperSearch Beyond What Can Be Taught: Evolving the Knowledge Boundary in Agentic Visual Generation
Related to (incoming)
contributed_to (incoming)
personhiyougapersonBUAADreamerpersonKuangdd01personcodemayqpersonfrozenleavespersonmarko1616personjiaqiw09personisLinXupersonZeyi-LinpersonMengqingCaopersonhzhaoypersontangeflypersonxvxuopoppersonLedzypersonBugmak3rKpersonerictang000personwangxingjun778personkhazicpersonmMrBunpersontastelikefeetpersonjohnnynunezpersonsunyi0505personJimmyPeilinLipersonjohannhartmannpersonshing100
Covers (incoming)
Related across the graph
paperFrom RGB Generation to Dense Field Readout: Pixel-Space Dense Prediction with Text-to-Image ModelspaperScalable Visual Pretraining for Language IntelligencetutorialFine-tune a small model on your own datapaperPanoWorld: Real-World Panoramic GenerationpersonMengqingCaopaperDynaKRAG: A Unified Framework for Learnable Evidence Control in Multi-Hop Retrieval-Augmented GenerationpaperGatedLinear: Adaptive Routing of Complementary Linear Bases for Time Series ForecastingpaperWat3R: Underwater 3D Geometry Learning without AnnotationspaperTexture++: Elevating 3D Asset Texture Resolution with a Region-Aware Diffusion ModelpaperWavefront Parallelization for Efficient Learned Image CompressionpaperLocalize, Then Reason: Visual Latent Structural Reasoning for Molecular Properties and EditsnewsOpenAI and Broadcom announce chip designed for LLM inference at scalepersonjohannhartmannpersontangeflypaperMonoIR-RS: Infrared Remote Sensing Vision-Language Learning with CLIP and VLM AdaptationpaperLongVQUBench: Benchmarking Long-Term Video Quality Understanding of Vision-Language ModelspaperCPI-Bench: A Comprehensive,Practical and Intelligent Benchmark for Real-World Image Editingpersonfrozenleavesmodelmeta-llama/Llama-3.1-8B-InstructpaperSynthetic-to-Real Translation for Class-Agnostic Motion PredictionpaperLLM Agents for Deliberative Collaboration: A Study on Joint Decision Making Under Partial ObservabilitypaperSMetric: Rethink LLM Scheduling for Serving Agents with Balanced Session-centric SchedulingpersonBugmak3rKpaperDominoTree: Conditional Tree-Structured Drafting with Domino for Speculative DecodingpaperPost-Training Shifts Confidence: A Three-Stage Analysis of How SFT, RL, and OPD Shape Pre-, Intra-, and Post-CoT CalibrationpersonZeyi-LinnewsOpenAI and Broadcom unveil LLM-optimized inference chippaperMM-IssueLoc: A Controlled Benchmark for Evaluating Visual Evidence in Multimodal Repository-Level Issue LocalizationpaperStable Self-Modulating Quantum Fast-Weight Programmers with Bounded Memory GatespaperReCal3R: Reliability-Calibrated Learning Rates for Streaming 3D ReconstructionnewsFine-Tuning Tool-Calling LLMs: A Complete Guide Using XYZ-Aquila-SFT and Qwen3 - MarkTechPostpaperFlash EQ-Linear: Accelerating Equivariant Linear Layers via Group-wise Discrete Fourier TransformnewsBeyond LoRA: Can you beat the most popular fine-tuning technique?personBUAADreamerpaperCan LLMs Write Reliable Rubrics? A Meta-Evaluation for Experiment ReproductionpersonxvxuopoppaperMulti-Modal, Multi-Environment Machine Teaching for Robust Reward LearningpersonkhazicpaperWeakly-Supervised RGB-D Salient Object Detection via SAM-driven Pseudo Annotation and State Space Interaction-based DiffusionpersonhzhaoypaperSIVA-RL: Sensitivity-Invariance Visual Alignment for Multimodal Reinforcement LearningpaperMemOps: Benchmarking Lifecycle Memory Operations in Long-Horizon Conversationspersonhiyougapersonerictang000paperOn Success and Simplicity: A Second Look at Transferable Vision-Language Attack PipelinepersonJimmyPeilinLipersonisLinXupaperVGOcc: Learning Visual-Geometric Gaussians for Vision-Centric 3D Driving Occupancy PredictionpaperScaling Behavior Foundation Model for Humanoid Robotspersontastelikefeetpersonsunyi0505modelmeta-llama/Meta-Llama-3-8BpaperGATE-3D: Geometry-Aware Test-time Adaptive Reranking for Open-Set 3D Shape RetrievalpaperFrom Style Replication to Style Exploration: Enabling Art Style Exploration with Analyze-Experiment-Resituate FrameworkpaperEAGLE-360: Embodied Active Global-to-Local Exploration in 360$^\circ$modeldeepseek-ai/DeepSeek-R1personmMrBunpaperChartGenEval: Corruption-Tested Multi-Dimensional Feedback for Rhythm-Game Chart GenerationpaperRethinking Quantum Continual Learning with Quantum Fisher InformationpersonLedzymodelmeta-llama/Meta-Llama-3-8B-Instructmodelmeta-llama/Llama-2-7bpaperPPL-Factory: Task-Aware and Budget-Aware Data Selection from Language Modeling to ReasoningpersoncodemayqpaperSparse Evidence Can Suffice: Agentic Evidence Seeking for Multimodal Video Misinformation DetectionpersonjohnnynunezpaperBeyond Single Expert: Harmonizing Diverse Visual Priors in MLLMs for Spatial UnderstandingpaperCross-Coordinate Correspondence Pruning for Image-to-Point Cloud RegistrationpaperTowards Enhancing 3D Spatial Reasoning in Medical Multimodal Large Language ModelspaperWhat Transfers Under Source Shift? Definitions, Examples, and Fine-Tuning for Climate Disclosure Classificationpersonjiaqiw09personmarko1616tutorialEvaluate a model properlypersonKuangdd01personwangxingjun778paperCR-Refiner: An Object-Centric Optimal Transport Reranker for Edit-Conditioned 3D Scene RetrievalpaperAutoregressive B-Rep Shape Generation with Parametric SurfacespaperDepthART: Scaling Foundation Monocular Depth to Tiny ModelspaperHumanTracker: Towards Comprehensive and Human-Aligned Motion Tracking Benchmarkmodelmeta-llama/Llama-2-7b-chat-hfpaperSNM-VFI: Symmetric Nonlinear Motion-Guided Generative Video Frame Interpolationmodelmeta-llama/Llama-3.3-70B-InstructpaperSearch Beyond What Can Be Taught: Evolving the Knowledge Boundary in Agentic Visual Generationpersonshing100paperMedGame: Storytelling Gamification Empowered by Large Language Models for Medical Education
