Skip to main content
Angestrom home
SearchPapersModelsLive AIIntelligence
Search⌕⌘K
EnterprisePricingSign in

Stay Ahead in the AI Revolution

Weekly digest — EPI pulse, top intelligence, fresh lineage. Free, no account.

Follow Angestrom
Global source network
Synced every 5 minutes

Continuous sync from primary AI sources — indexed, enriched, and queryable in real time.

arXivHugging FaceGitHubOpenAIAnthropicDeepMindReutersBBC TechHacker NewsReddit MLVerified feedsFunding
ANGESTROM

The Intelligence Layer of Humanity. Everything AI. All in One Place.

Angestrom connects every piece of the AI ecosystem — data, models, research, companies, tools, and people.

info@angestrom.comwww.angestrom.comLucknow, Uttar Pradesh, India

Product

  • AI Search
  • AI Models
  • Research Papers
  • Companies
  • News & Events
  • GitHub Explorer
  • APIs & Tools
  • Datasets
  • Benchmarks
  • Model lifecycle
  • Funding graph
  • Contributors
  • AI Agents

Resources

  • Weekly digest
  • Documentation
  • Tutorials
  • Guides
  • News
  • Help / Start
  • Community

Company

  • About
  • Contact
  • Privacy Policy
  • Terms of Service
  • Acceptable Use

Enterprise

  • Pricing
  • Workspace
  • Contact Sales

Developer

  • Developer Hub
  • API docs
  • GitHub

Learn

  • Learning Academy
  • Roadmaps
  • Glossary
  • AI for Beginners

Popular Topics

Loading topics…
View All Topics →
© 2026 Angestrom Intelligence Private Limited. All rights reserved.
English
Theme
Angestrom home
SearchPapersModelsLive AIIntelligence
Search⌕⌘K
EnterprisePricingSign in
  1. Home
  2. /Repositories
  3. /hiyouga/LlamaFactory
Read original ↗
repoGitHubTrust 82 · PrimaryPublished 1mo agoLive · 24d ago

hiyouga/LlamaFactory

Unified Efficient Fine-Tuning of 100+ LLMs & VLMs (ACL 2024)

Lineage graph

Paper → model → repo connections mined from source citations (Tier-1 exact match).

Why these links exist

Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.

  • PossiblePossibly related (embedding) · 53%Fine-tune a small model on your own data →
  • PossiblePossibly related (embedding) · 53%Beyond LoRA: Can you beat the most popular fine-tuning technique? →
  • PossiblePossibly related (embedding) · 50%OpenAI and Broadcom announce chip designed for LLM inference at scale →
  • PossiblePossibly related (embedding) · 48%Evaluate a model properly →
  • PossiblePossibly related (embedding) · 47%OpenAI and Broadcom unveil LLM-optimized inference chip →
  • FuzzyOverlapping authors or contributors · 62%PanoWorld: Real-World Panoramic Generation →

    “Shared author/contributor keys: lin”

  • FuzzyOverlapping authors or contributors · 62%DynaKRAG: A Unified Framework for Learnable Evidence Control in Multi-Hop Retrieval-Augmented Generation →

    “Shared author/contributor keys: lin”

  • FuzzyOverlapping authors or contributors · 62%GatedLinear: Adaptive Routing of Complementary Linear Bases for Time Series Forecasting →

    “Shared author/contributor keys: lin”

Related to

tutorialFine-tune a small model on your own datatutorialEvaluate a model properly

Covers

newsBeyond LoRA: Can you beat the most popular fine-tuning technique?newsOpenAI and Broadcom announce chip designed for LLM inference at scalenewsOpenAI and Broadcom unveil LLM-optimized inference chip

Implements

paperPanoWorld: Real-World Panoramic GenerationpaperDynaKRAG: A Unified Framework for Learnable Evidence Control in Multi-Hop Retrieval-Augmented GenerationpaperGatedLinear: Adaptive Routing of Complementary Linear Bases for Time Series ForecastingpaperWat3R: Underwater 3D Geometry Learning without AnnotationspaperLongVQUBench: Benchmarking Long-Term Video Quality Understanding of Vision-Language ModelspaperSynthetic-to-Real Translation for Class-Agnostic Motion PredictionpaperDominoTree: Conditional Tree-Structured Drafting with Domino for Speculative DecodingpaperMM-IssueLoc: A Controlled Benchmark for Evaluating Visual Evidence in Multimodal Repository-Level Issue LocalizationpaperScaling Behavior Foundation Model for Humanoid RobotspaperBeyond Single Expert: Harmonizing Diverse Visual Priors in MLLMs for Spatial UnderstandingpaperWeakly-Supervised RGB-D Salient Object Detection via SAM-driven Pseudo Annotation and State Space Interaction-based DiffusionpaperOn Success and Simplicity: A Second Look at Transferable Vision-Language Attack PipelinepaperDepthART: Scaling Foundation Monocular Depth to Tiny ModelspaperCross-Coordinate Correspondence Pruning for Image-to-Point Cloud RegistrationpaperAutoregressive B-Rep Shape Generation with Parametric SurfacespaperMulti-Modal, Multi-Environment Machine Teaching for Robust Reward LearningpaperCR-Refiner: An Object-Centric Optimal Transport Reranker for Edit-Conditioned 3D Scene RetrievalpaperGATE-3D: Geometry-Aware Test-time Adaptive Reranking for Open-Set 3D Shape RetrievalpaperWavefront Parallelization for Efficient Learned Image CompressionpaperSearch Beyond What Can Be Taught: Evolving the Knowledge Boundary in Agentic Visual GenerationpaperFrom RGB Generation to Dense Field Readout: Pixel-Space Dense Prediction with Text-to-Image ModelspaperTowards Enhancing 3D Spatial Reasoning in Medical Multimodal Large Language ModelspaperMonoIR-RS: Infrared Remote Sensing Vision-Language Learning with CLIP and VLM AdaptationpaperPost-Training Shifts Confidence: A Three-Stage Analysis of How SFT, RL, and OPD Shape Pre-, Intra-, and Post-CoT CalibrationpaperStable Self-Modulating Quantum Fast-Weight Programmers with Bounded Memory GatespaperSparse Evidence Can Suffice: Agentic Evidence Seeking for Multimodal Video Misinformation DetectionpaperPPL-Factory: Task-Aware and Budget-Aware Data Selection from Language Modeling to ReasoningpaperVGOcc: Learning Visual-Geometric Gaussians for Vision-Centric 3D Driving Occupancy PredictionpaperWhat Transfers Under Source Shift? Definitions, Examples, and Fine-Tuning for Climate Disclosure ClassificationpaperMemOps: Benchmarking Lifecycle Memory Operations in Long-Horizon ConversationspaperTexture++: Elevating 3D Asset Texture Resolution with a Region-Aware Diffusion ModelpaperMedGame: Storytelling Gamification Empowered by Large Language Models for Medical EducationpaperLocalize, Then Reason: Visual Latent Structural Reasoning for Molecular Properties and EditspaperHumanTracker: Towards Comprehensive and Human-Aligned Motion Tracking BenchmarkpaperSNM-VFI: Symmetric Nonlinear Motion-Guided Generative Video Frame InterpolationpaperDiagnosing Dense Same-Class Attribute Misbinding in Large Vision-Language ModelspaperUniDot: A Unified Network for Sequence Modeling and Feature Interaction in Large-scale RecommendationpaperSTAGE: Controlled Objective Admission for Multi-Preference LLM AlignmentpaperUnlocking the Potential of Image Editing via Concept Scaling and Dense SupervisionpaperReCal3R: Reliability-Calibrated Learning Rates for Streaming 3D ReconstructionpaperLLM Agents for Deliberative Collaboration: A Study on Joint Decision Making Under Partial ObservabilitypaperSMetric: Rethink LLM Scheduling for Serving Agents with Balanced Session-centric SchedulingpaperCan LLMs Write Reliable Rubrics? A Meta-Evaluation for Experiment ReproductionpaperSIVA-RL: Sensitivity-Invariance Visual Alignment for Multimodal Reinforcement LearningpaperChartGenEval: Corruption-Tested Multi-Dimensional Feedback for Rhythm-Game Chart GenerationpaperEAGLE-360: Embodied Active Global-to-Local Exploration in 360$^\circ$paperRethinking Quantum Continual Learning with Quantum Fisher InformationpaperFlash EQ-Linear: Accelerating Equivariant Linear Layers via Group-wise Discrete Fourier TransformpaperFrom Style Replication to Style Exploration: Enabling Art Style Exploration with Analyze-Experiment-Resituate FrameworkpaperCPI-Bench: A Comprehensive,Practical and Intelligent Benchmark for Real-World Image EditingpaperScalable Visual Pretraining for Language Intelligence

Related to (incoming)

modeldeepseek-ai/DeepSeek-R1modelmeta-llama/Meta-Llama-3-8Bmodelmeta-llama/Llama-3.1-8B-Instructmodelmeta-llama/Llama-2-7b-chat-hfmodelmeta-llama/Meta-Llama-3-8B-Instructmodelmeta-llama/Llama-2-7bmodelmeta-llama/Llama-3.3-70B-Instruct

contributed_to (incoming)

personhiyougapersonBUAADreamerpersonKuangdd01personcodemayqpersonfrozenleavespersonmarko1616personjiaqiw09personisLinXupersonZeyi-LinpersonMengqingCaopersonhzhaoypersontangeflypersonxvxuopoppersonLedzypersonBugmak3rKpersonerictang000personwangxingjun778personkhazicpersonmMrBunpersontastelikefeetpersonjohnnynunezpersonsunyi0505personJimmyPeilinLipersonjohannhartmannpersonshing100

Covers (incoming)

newsFine-Tuning Tool-Calling LLMs: A Complete Guide Using XYZ-Aquila-SFT and Qwen3 - MarkTechPost

Related across the graph

paperFrom RGB Generation to Dense Field Readout: Pixel-Space Dense Prediction with Text-to-Image ModelspaperScalable Visual Pretraining for Language IntelligencetutorialFine-tune a small model on your own datapaperDiagnosing Dense Same-Class Attribute Misbinding in Large Vision-Language ModelspaperPanoWorld: Real-World Panoramic GenerationpersonMengqingCaopaperDynaKRAG: A Unified Framework for Learnable Evidence Control in Multi-Hop Retrieval-Augmented GenerationpaperGatedLinear: Adaptive Routing of Complementary Linear Bases for Time Series ForecastingpaperWat3R: Underwater 3D Geometry Learning without AnnotationspaperTexture++: Elevating 3D Asset Texture Resolution with a Region-Aware Diffusion ModelpaperWavefront Parallelization for Efficient Learned Image CompressionpaperLocalize, Then Reason: Visual Latent Structural Reasoning for Molecular Properties and EditsnewsOpenAI and Broadcom announce chip designed for LLM inference at scalepaperUniDot: A Unified Network for Sequence Modeling and Feature Interaction in Large-scale RecommendationpersonjohannhartmannpersontangeflypaperMonoIR-RS: Infrared Remote Sensing Vision-Language Learning with CLIP and VLM AdaptationpaperLongVQUBench: Benchmarking Long-Term Video Quality Understanding of Vision-Language ModelspaperCPI-Bench: A Comprehensive,Practical and Intelligent Benchmark for Real-World Image Editingpersonfrozenleavesmodelmeta-llama/Llama-3.1-8B-InstructpaperSynthetic-to-Real Translation for Class-Agnostic Motion PredictionpaperLLM Agents for Deliberative Collaboration: A Study on Joint Decision Making Under Partial ObservabilitypaperSMetric: Rethink LLM Scheduling for Serving Agents with Balanced Session-centric SchedulingpersonBugmak3rKpaperDominoTree: Conditional Tree-Structured Drafting with Domino for Speculative DecodingpaperPost-Training Shifts Confidence: A Three-Stage Analysis of How SFT, RL, and OPD Shape Pre-, Intra-, and Post-CoT CalibrationpersonZeyi-LinnewsOpenAI and Broadcom unveil LLM-optimized inference chippaperMM-IssueLoc: A Controlled Benchmark for Evaluating Visual Evidence in Multimodal Repository-Level Issue LocalizationpaperStable Self-Modulating Quantum Fast-Weight Programmers with Bounded Memory GatespaperReCal3R: Reliability-Calibrated Learning Rates for Streaming 3D ReconstructionnewsFine-Tuning Tool-Calling LLMs: A Complete Guide Using XYZ-Aquila-SFT and Qwen3 - MarkTechPostpaperFlash EQ-Linear: Accelerating Equivariant Linear Layers via Group-wise Discrete Fourier TransformnewsBeyond LoRA: Can you beat the most popular fine-tuning technique?personBUAADreamerpaperCan LLMs Write Reliable Rubrics? A Meta-Evaluation for Experiment ReproductionpersonxvxuopoppaperMulti-Modal, Multi-Environment Machine Teaching for Robust Reward LearningpersonkhazicpaperWeakly-Supervised RGB-D Salient Object Detection via SAM-driven Pseudo Annotation and State Space Interaction-based DiffusionpersonhzhaoypaperSIVA-RL: Sensitivity-Invariance Visual Alignment for Multimodal Reinforcement LearningpaperMemOps: Benchmarking Lifecycle Memory Operations in Long-Horizon Conversationspersonhiyougapersonerictang000paperOn Success and Simplicity: A Second Look at Transferable Vision-Language Attack PipelinepersonJimmyPeilinLipersonisLinXupaperVGOcc: Learning Visual-Geometric Gaussians for Vision-Centric 3D Driving Occupancy PredictionpaperScaling Behavior Foundation Model for Humanoid Robotspersontastelikefeetpersonsunyi0505modelmeta-llama/Meta-Llama-3-8BpaperGATE-3D: Geometry-Aware Test-time Adaptive Reranking for Open-Set 3D Shape RetrievalpaperFrom Style Replication to Style Exploration: Enabling Art Style Exploration with Analyze-Experiment-Resituate FrameworkpaperEAGLE-360: Embodied Active Global-to-Local Exploration in 360$^\circ$modeldeepseek-ai/DeepSeek-R1personmMrBunpaperChartGenEval: Corruption-Tested Multi-Dimensional Feedback for Rhythm-Game Chart GenerationpaperRethinking Quantum Continual Learning with Quantum Fisher InformationpersonLedzymodelmeta-llama/Meta-Llama-3-8B-Instructmodelmeta-llama/Llama-2-7bpaperSTAGE: Controlled Objective Admission for Multi-Preference LLM AlignmentpaperPPL-Factory: Task-Aware and Budget-Aware Data Selection from Language Modeling to ReasoningpersoncodemayqpaperSparse Evidence Can Suffice: Agentic Evidence Seeking for Multimodal Video Misinformation DetectionpersonjohnnynunezpaperBeyond Single Expert: Harmonizing Diverse Visual Priors in MLLMs for Spatial UnderstandingpaperCross-Coordinate Correspondence Pruning for Image-to-Point Cloud RegistrationpaperTowards Enhancing 3D Spatial Reasoning in Medical Multimodal Large Language ModelspaperUnlocking the Potential of Image Editing via Concept Scaling and Dense SupervisionpaperWhat Transfers Under Source Shift? Definitions, Examples, and Fine-Tuning for Climate Disclosure Classificationpersonjiaqiw09personmarko1616tutorialEvaluate a model properlypersonKuangdd01personwangxingjun778paperCR-Refiner: An Object-Centric Optimal Transport Reranker for Edit-Conditioned 3D Scene RetrievalpaperAutoregressive B-Rep Shape Generation with Parametric SurfacespaperDepthART: Scaling Foundation Monocular Depth to Tiny ModelspaperHumanTracker: Towards Comprehensive and Human-Aligned Motion Tracking Benchmarkmodelmeta-llama/Llama-2-7b-chat-hfpaperSNM-VFI: Symmetric Nonlinear Motion-Guided Generative Video Frame Interpolationmodelmeta-llama/Llama-3.3-70B-InstructpaperSearch Beyond What Can Be Taught: Evolving the Knowledge Boundary in Agentic Visual Generationpersonshing100paperMedGame: Storytelling Gamification Empowered by Large Language Models for Medical Education
Knowledge path·PFrom RGB Generation to Dense Field Readout: Pixel-Space Dense Prediction with Text-to-Image Models→PScalable Visual Pretraining for Language Intelligence→LFine-tune a small model on your own data→Rhiyouga/LlamaFactory

Topics

agentaideepseekfine-tuninggemmagptinstruction-tuninglarge-language-modelsllamallama3

Explore

Search similar →Knowledge graph →All repos →Full intelligence feed →
Maintenance82
RIS100GitHub verified
Graph trust82Primary
Graph score73480