Skip to main content
Angestrom home
SearchPapersModelsLive AIIntelligence
Search⌕⌘K
EnterprisePricingSign in

Stay Ahead in the AI Revolution

Weekly digest — EPI pulse, top intelligence, fresh lineage. Free, no account.

Follow Angestrom
Global source network
Synced every 5 minutes

Continuous sync from primary AI sources — indexed, enriched, and queryable in real time.

arXivHugging FaceGitHubOpenAIAnthropicDeepMindReutersBBC TechHacker NewsReddit MLVerified feedsFunding
ANGESTROM

The Intelligence Layer of Humanity. Everything AI. All in One Place.

Angestrom connects every piece of the AI ecosystem — data, models, research, companies, tools, and people.

info@angestrom.comwww.angestrom.comLucknow, Uttar Pradesh, India

Product

  • AI Search
  • AI Models
  • Research Papers
  • Companies
  • News & Events
  • GitHub Explorer
  • APIs & Tools
  • Datasets
  • Benchmarks
  • Model lifecycle
  • Funding graph
  • Contributors
  • AI Agents

Resources

  • Weekly digest
  • Documentation
  • Tutorials
  • Guides
  • News
  • Help / Start
  • Community

Company

  • About
  • Contact
  • Privacy Policy
  • Terms of Service
  • Acceptable Use

Enterprise

  • Pricing
  • Workspace
  • Contact Sales

Developer

  • Developer Hub
  • API docs
  • GitHub

Learn

  • Learning Academy
  • Roadmaps
  • Glossary
  • AI for Beginners

Popular Topics

Loading topics…
View All Topics →
© 2026 Angestrom Intelligence Private Limited. All rights reserved.
English
Theme
Angestrom home
SearchPapersModelsLive AIIntelligence
Search⌕⌘K
EnterprisePricingSign in
  1. Home
  2. /Repositories
  3. /hiyouga/LlamaFactory
Read original ↗
repoGitHubTrust 82 · PrimaryPublished 1mo agoLive · 24d ago

hiyouga/LlamaFactory

Unified Efficient Fine-Tuning of 100+ LLMs & VLMs (ACL 2024)

Lineage graph

Paper → model → repo connections mined from source citations (Tier-1 exact match).

Why these links exist

Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.

  • PossiblePossibly related (embedding) · 53%Fine-tune a small model on your own data →
  • PossiblePossibly related (embedding) · 53%Beyond LoRA: Can you beat the most popular fine-tuning technique? →
  • PossiblePossibly related (embedding) · 50%OpenAI and Broadcom announce chip designed for LLM inference at scale →
  • PossiblePossibly related (embedding) · 48%Evaluate a model properly →
  • PossiblePossibly related (embedding) · 47%OpenAI and Broadcom unveil LLM-optimized inference chip →
  • FuzzyOverlapping authors or contributors · 62%From RGB Generation to Dense Field Readout: Pixel-Space Dense Prediction with Text-to-Image Models →

    “Shared author/contributor keys: lin”

  • FuzzyOverlapping authors or contributors · 62%Towards Enhancing 3D Spatial Reasoning in Medical Multimodal Large Language Models →

    “Shared author/contributor keys: lin”

  • FuzzyOverlapping authors or contributors · 62%MonoIR-RS: Infrared Remote Sensing Vision-Language Learning with CLIP and VLM Adaptation →

    “Shared author/contributor keys: lin”

Related to

tutorialFine-tune a small model on your own datatutorialEvaluate a model properly

Covers

newsBeyond LoRA: Can you beat the most popular fine-tuning technique?newsOpenAI and Broadcom announce chip designed for LLM inference at scalenewsOpenAI and Broadcom unveil LLM-optimized inference chip

Implements

paperFrom RGB Generation to Dense Field Readout: Pixel-Space Dense Prediction with Text-to-Image ModelspaperTowards Enhancing 3D Spatial Reasoning in Medical Multimodal Large Language ModelspaperMonoIR-RS: Infrared Remote Sensing Vision-Language Learning with CLIP and VLM AdaptationpaperPost-Training Shifts Confidence: A Three-Stage Analysis of How SFT, RL, and OPD Shape Pre-, Intra-, and Post-CoT CalibrationpaperPanoWorld: Real-World Panoramic GenerationpaperDynaKRAG: A Unified Framework for Learnable Evidence Control in Multi-Hop Retrieval-Augmented GenerationpaperGatedLinear: Adaptive Routing of Complementary Linear Bases for Time Series ForecastingpaperWat3R: Underwater 3D Geometry Learning without AnnotationspaperLongVQUBench: Benchmarking Long-Term Video Quality Understanding of Vision-Language ModelspaperSynthetic-to-Real Translation for Class-Agnostic Motion PredictionpaperDominoTree: Conditional Tree-Structured Drafting with Domino for Speculative DecodingpaperStable Self-Modulating Quantum Fast-Weight Programmers with Bounded Memory GatespaperReCal3R: Reliability-Calibrated Learning Rates for Streaming 3D ReconstructionpaperLLM Agents for Deliberative Collaboration: A Study on Joint Decision Making Under Partial ObservabilitypaperSMetric: Rethink LLM Scheduling for Serving Agents with Balanced Session-centric SchedulingpaperCan LLMs Write Reliable Rubrics? A Meta-Evaluation for Experiment ReproductionpaperSIVA-RL: Sensitivity-Invariance Visual Alignment for Multimodal Reinforcement LearningpaperMM-IssueLoc: A Controlled Benchmark for Evaluating Visual Evidence in Multimodal Repository-Level Issue LocalizationpaperScaling Behavior Foundation Model for Humanoid RobotspaperBeyond Single Expert: Harmonizing Diverse Visual Priors in MLLMs for Spatial UnderstandingpaperWeakly-Supervised RGB-D Salient Object Detection via SAM-driven Pseudo Annotation and State Space Interaction-based DiffusionpaperOn Success and Simplicity: A Second Look at Transferable Vision-Language Attack PipelinepaperChartGenEval: Corruption-Tested Multi-Dimensional Feedback for Rhythm-Game Chart GenerationpaperEAGLE-360: Embodied Active Global-to-Local Exploration in 360$^\circ$paperRethinking Quantum Continual Learning with Quantum Fisher InformationpaperDepthART: Scaling Foundation Monocular Depth to Tiny ModelspaperCross-Coordinate Correspondence Pruning for Image-to-Point Cloud RegistrationpaperAutoregressive B-Rep Shape Generation with Parametric SurfacespaperSparse Evidence Can Suffice: Agentic Evidence Seeking for Multimodal Video Misinformation DetectionpaperPPL-Factory: Task-Aware and Budget-Aware Data Selection from Language Modeling to ReasoningpaperVGOcc: Learning Visual-Geometric Gaussians for Vision-Centric 3D Driving Occupancy PredictionpaperWhat Transfers Under Source Shift? Definitions, Examples, and Fine-Tuning for Climate Disclosure ClassificationpaperMulti-Modal, Multi-Environment Machine Teaching for Robust Reward LearningpaperMemOps: Benchmarking Lifecycle Memory Operations in Long-Horizon ConversationspaperCR-Refiner: An Object-Centric Optimal Transport Reranker for Edit-Conditioned 3D Scene RetrievalpaperGATE-3D: Geometry-Aware Test-time Adaptive Reranking for Open-Set 3D Shape RetrievalpaperWavefront Parallelization for Efficient Learned Image CompressionpaperFlash EQ-Linear: Accelerating Equivariant Linear Layers via Group-wise Discrete Fourier TransformpaperTexture++: Elevating 3D Asset Texture Resolution with a Region-Aware Diffusion ModelpaperMedGame: Storytelling Gamification Empowered by Large Language Models for Medical EducationpaperLocalize, Then Reason: Visual Latent Structural Reasoning for Molecular Properties and EditspaperHumanTracker: Towards Comprehensive and Human-Aligned Motion Tracking BenchmarkpaperSNM-VFI: Symmetric Nonlinear Motion-Guided Generative Video Frame InterpolationpaperFrom Style Replication to Style Exploration: Enabling Art Style Exploration with Analyze-Experiment-Resituate FrameworkpaperCPI-Bench: A Comprehensive,Practical and Intelligent Benchmark for Real-World Image EditingpaperScalable Visual Pretraining for Language IntelligencepaperSearch Beyond What Can Be Taught: Evolving the Knowledge Boundary in Agentic Visual Generation

Related to (incoming)

modeldeepseek-ai/DeepSeek-R1modelmeta-llama/Meta-Llama-3-8Bmodelmeta-llama/Llama-3.1-8B-Instructmodelmeta-llama/Llama-2-7b-chat-hfmodelmeta-llama/Meta-Llama-3-8B-Instructmodelmeta-llama/Llama-2-7bmodelmeta-llama/Llama-3.3-70B-Instruct

contributed_to (incoming)

personhiyougapersonBUAADreamerpersonKuangdd01personcodemayqpersonfrozenleavespersonmarko1616personjiaqiw09personisLinXupersonZeyi-LinpersonMengqingCaopersonhzhaoypersontangeflypersonxvxuopoppersonLedzypersonBugmak3rKpersonerictang000personwangxingjun778personkhazicpersonmMrBunpersontastelikefeetpersonjohnnynunezpersonsunyi0505personJimmyPeilinLipersonjohannhartmannpersonshing100

Covers (incoming)

newsFine-Tuning Tool-Calling LLMs: A Complete Guide Using XYZ-Aquila-SFT and Qwen3 - MarkTechPost

Related across the graph

paperFrom RGB Generation to Dense Field Readout: Pixel-Space Dense Prediction with Text-to-Image ModelspaperScalable Visual Pretraining for Language IntelligencetutorialFine-tune a small model on your own datapaperPanoWorld: Real-World Panoramic GenerationpersonMengqingCaopaperDynaKRAG: A Unified Framework for Learnable Evidence Control in Multi-Hop Retrieval-Augmented GenerationpaperGatedLinear: Adaptive Routing of Complementary Linear Bases for Time Series ForecastingpaperWat3R: Underwater 3D Geometry Learning without AnnotationspaperTexture++: Elevating 3D Asset Texture Resolution with a Region-Aware Diffusion ModelpaperWavefront Parallelization for Efficient Learned Image CompressionpaperLocalize, Then Reason: Visual Latent Structural Reasoning for Molecular Properties and EditsnewsOpenAI and Broadcom announce chip designed for LLM inference at scalepersonjohannhartmannpersontangeflypaperMonoIR-RS: Infrared Remote Sensing Vision-Language Learning with CLIP and VLM AdaptationpaperLongVQUBench: Benchmarking Long-Term Video Quality Understanding of Vision-Language ModelspaperCPI-Bench: A Comprehensive,Practical and Intelligent Benchmark for Real-World Image Editingpersonfrozenleavesmodelmeta-llama/Llama-3.1-8B-InstructpaperSynthetic-to-Real Translation for Class-Agnostic Motion PredictionpaperLLM Agents for Deliberative Collaboration: A Study on Joint Decision Making Under Partial ObservabilitypaperSMetric: Rethink LLM Scheduling for Serving Agents with Balanced Session-centric SchedulingpersonBugmak3rKpaperDominoTree: Conditional Tree-Structured Drafting with Domino for Speculative DecodingpaperPost-Training Shifts Confidence: A Three-Stage Analysis of How SFT, RL, and OPD Shape Pre-, Intra-, and Post-CoT CalibrationpersonZeyi-LinnewsOpenAI and Broadcom unveil LLM-optimized inference chippaperMM-IssueLoc: A Controlled Benchmark for Evaluating Visual Evidence in Multimodal Repository-Level Issue LocalizationpaperStable Self-Modulating Quantum Fast-Weight Programmers with Bounded Memory GatespaperReCal3R: Reliability-Calibrated Learning Rates for Streaming 3D ReconstructionnewsFine-Tuning Tool-Calling LLMs: A Complete Guide Using XYZ-Aquila-SFT and Qwen3 - MarkTechPostpaperFlash EQ-Linear: Accelerating Equivariant Linear Layers via Group-wise Discrete Fourier TransformnewsBeyond LoRA: Can you beat the most popular fine-tuning technique?personBUAADreamerpaperCan LLMs Write Reliable Rubrics? A Meta-Evaluation for Experiment ReproductionpersonxvxuopoppaperMulti-Modal, Multi-Environment Machine Teaching for Robust Reward LearningpersonkhazicpaperWeakly-Supervised RGB-D Salient Object Detection via SAM-driven Pseudo Annotation and State Space Interaction-based DiffusionpersonhzhaoypaperSIVA-RL: Sensitivity-Invariance Visual Alignment for Multimodal Reinforcement LearningpaperMemOps: Benchmarking Lifecycle Memory Operations in Long-Horizon Conversationspersonhiyougapersonerictang000paperOn Success and Simplicity: A Second Look at Transferable Vision-Language Attack PipelinepersonJimmyPeilinLipersonisLinXupaperVGOcc: Learning Visual-Geometric Gaussians for Vision-Centric 3D Driving Occupancy PredictionpaperScaling Behavior Foundation Model for Humanoid Robotspersontastelikefeetpersonsunyi0505modelmeta-llama/Meta-Llama-3-8BpaperGATE-3D: Geometry-Aware Test-time Adaptive Reranking for Open-Set 3D Shape RetrievalpaperFrom Style Replication to Style Exploration: Enabling Art Style Exploration with Analyze-Experiment-Resituate FrameworkpaperEAGLE-360: Embodied Active Global-to-Local Exploration in 360$^\circ$modeldeepseek-ai/DeepSeek-R1personmMrBunpaperChartGenEval: Corruption-Tested Multi-Dimensional Feedback for Rhythm-Game Chart GenerationpaperRethinking Quantum Continual Learning with Quantum Fisher InformationpersonLedzymodelmeta-llama/Meta-Llama-3-8B-Instructmodelmeta-llama/Llama-2-7bpaperPPL-Factory: Task-Aware and Budget-Aware Data Selection from Language Modeling to ReasoningpersoncodemayqpaperSparse Evidence Can Suffice: Agentic Evidence Seeking for Multimodal Video Misinformation DetectionpersonjohnnynunezpaperBeyond Single Expert: Harmonizing Diverse Visual Priors in MLLMs for Spatial UnderstandingpaperCross-Coordinate Correspondence Pruning for Image-to-Point Cloud RegistrationpaperTowards Enhancing 3D Spatial Reasoning in Medical Multimodal Large Language ModelspaperWhat Transfers Under Source Shift? Definitions, Examples, and Fine-Tuning for Climate Disclosure Classificationpersonjiaqiw09personmarko1616tutorialEvaluate a model properlypersonKuangdd01personwangxingjun778paperCR-Refiner: An Object-Centric Optimal Transport Reranker for Edit-Conditioned 3D Scene RetrievalpaperAutoregressive B-Rep Shape Generation with Parametric SurfacespaperDepthART: Scaling Foundation Monocular Depth to Tiny ModelspaperHumanTracker: Towards Comprehensive and Human-Aligned Motion Tracking Benchmarkmodelmeta-llama/Llama-2-7b-chat-hfpaperSNM-VFI: Symmetric Nonlinear Motion-Guided Generative Video Frame Interpolationmodelmeta-llama/Llama-3.3-70B-InstructpaperSearch Beyond What Can Be Taught: Evolving the Knowledge Boundary in Agentic Visual Generationpersonshing100paperMedGame: Storytelling Gamification Empowered by Large Language Models for Medical Education
Knowledge path·PFrom RGB Generation to Dense Field Readout: Pixel-Space Dense Prediction with Text-to-Image Models→PScalable Visual Pretraining for Language Intelligence→LFine-tune a small model on your own data→Rhiyouga/LlamaFactory

Topics

agentaideepseekfine-tuninggemmagptinstruction-tuninglarge-language-modelsllamallama3

Explore

Search similar →Knowledge graph →All repos →Full intelligence feed →
Maintenance82
RIS100GitHub verified
Graph trust82Primary
Graph score73480