repoGitHubTrust 82 Β· PrimaryPublished 1mo agoLive Β· 10h ago
huggingface/transformers
π€ Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and training.
Lineage graph
Paper β model β repo connections mined from source citations (Tier-1 exact match).
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it β so bad links are debuggable.
- LinkedLinked via dependency parse Β· 85%c4 β
βListed in pyproject.tomlβ
- LinkedLinked via dependency parse Β· 85%c901 β
βListed in pyproject.tomlβ
- LinkedLinked via dependency parse Β· 85%e β
βListed in pyproject.tomlβ
- LinkedLinked via dependency parse Β· 85%e501 β
βListed in pyproject.tomlβ
- LinkedLinked via dependency parse Β· 85%e741 β
βListed in pyproject.tomlβ
- LinkedLinked via dependency parse Β· 85%except β
βListed in pyproject.tomlβ
- LinkedLinked via dependency parse Β· 85%f β
βListed in pyproject.tomlβ
- LinkedLinked via dependency parse Β· 85%furb β
βListed in pyproject.tomlβ
Related to
Implements
paperGrokking in small transformerspaperInhibited Self-Attention: Sharpening Focus in Vision TransformerspaperAutomated Compliance Mapping in Cloud Security with Domain-Adapted Sentence TransformerspaperPAC-ACT: Post-training Actor-Critic for Action Chunking TransformerspaperFoveation-Guided Dynamic Token Selection for Robust and Efficient Vision TransformerspaperFrom Expressivity to Sample Complexity: Narrow Teachers for Transformers via C-RASPpaperQuasiMoTTo: Quasi-Monte Carlo Test-Time ScalingpaperInvariant Learning Dynamics of Transformers in Inductive Reasoning TaskspaperPost-Training Pruning for Diffusion TransformerspaperMobius Learning: Cyclic Depth Folding in TransformerspaperAppearance Pointers -- Multimodal Region Control of Diffusion TransformerspaperText Template Tokens Are Implicit Semantic Registers in Diffusion TransformerspaperELSAA: Efficient Low-Rank and Sparse Attention Approximation for Training TransformerspaperKaleido: Algorithm-Hardware Co-Design for Video Diffusion Transformers by Exploiting Latent Space CorrelationspaperMxGPS: Multiplex Graph Transformers for a Power Grid Foundation ModelpaperFlexViT: A Flexible FPGA-based Accelerator for Edge Vision TransformerspaperReview Residuals: Update-Conditioned Residual Gating for TransformerspaperUmm... With Transformers? Insights from Filled Pause Use across Four Slavic ParliamentspaperKroQuant: Kronecker-Structured Block Transforms for Efficient Post-Training Quantization of Diffusion Transformers
Related to (incoming)
modelmistralai/Mixtral-8x7B-Instruct-v0.1modelmicrosoft/phi-2modelzai-org/GLM-5.2modelmistralai/Mistral-7B-Instruct-v0.2modelopenai/whisper-large-v3-turbomodeldeepseek-ai/DeepSeek-V3modeldeepseek-ai/DeepSeek-R1modeldeepseek-ai/DeepSeek-V4-Promodeldeepseek-ai/Janus-Pro-7BpaperA Large-Scale Measurement of AI Bill of Materials Completeness in Hugging Face Modelsmodelsentence-transformers/all-MiniLM-L6-v2modeldeepseek-ai/DeepSeek-V3-0324modelBAAI/bge-m3modeldeepseek-ai/DeepSeek-OCR
Implements (incoming)
paperMultimodal Knowledge Edit-Scoped Generalization for Online Recursive MLLM EditingpaperTransformer Geometry Observatory TGO-II: Representational Similarity ObservatorypaperDeform360: A Massive Multi-view Visuotactile Dataset for Deformable World ModelspaperCAIRN: Cross-Room 3D Scene Understanding with Topology-Aware Large Multimodal ModelspaperMotion4Motion: Motion Transfer Across Subjects at InferencepaperANGLE: Angular Neural Generative Learning via EngressionpaperDo We Really Need Multimodal Emotion Language Models Larger Than 1B Parameters?paperXRFormer: Multiscale Tokenization for XRF Representation LearningpaperSVF-CR: Synchronized Visual-Facial Cross-Refinement for Multimodal Ambivalence and Hesitancy Recognition
contributed_to (incoming)
personydshiehpersonthomwolfpersonsguggerpersonLysandreJikpersonpatrickvonplatenpersongantepersonstas00personArthurZuckerpersonzucchini-nlppersonRocketknight1personjulien-cpersonCyrilvallezpersonyounesbelkadapersonNielsRoggepersonsshleiferpersonSunMarcpersonamyerobertspersonNarsilpersonstevhliupersonpatil-surajpersonVictorSanhpersoncyyeverpersonyonigozlanpersondependabot[bot]personsanchit-gandhi
Covers (incoming)
newsTransformers in Deep Learning: How Self-Attention Changed Modern AI - SnowflakenewsMicrosoft unveils MAI-Image-1, its first AI model that turns words into pictures - The Eastleigh VoicenewsAutoencoders: Learning Through Reconstruction - SnowflakenewsFLUX 3 - Real World Models: Towards Multimodal Flow Models as the Backbone of Visual Intelligence
Related across the graph
toolipersonyounesbelkadapersonsshleifertoolup031paperReview Residuals: Update-Conditioned Residual Gating for TransformerspaperAutomated Compliance Mapping in Cloud Security with Domain-Adapted Sentence TransformerstoolfurbpaperInhibited Self-Attention: Sharpening Focus in Vision TransformersmodelBAAI/bge-m3toolepersonNielsRoggepaperA Large-Scale Measurement of AI Bill of Materials Completeness in Hugging Face Modelspersonyonigozlantoole501modelzai-org/GLM-5.2modelmicrosoft/phi-2paperAppearance Pointers -- Multimodal Region Control of Diffusion TransformerspaperXRFormer: Multiscale Tokenization for XRF Representation LearningpaperPAC-ACT: Post-training Actor-Critic for Action Chunking Transformersmodeldeepseek-ai/DeepSeek-V3toolpie794paperFrom Expressivity to Sample Complexity: Narrow Teachers for Transformers via C-RASPpaperFoveation-Guided Dynamic Token Selection for Robust and Efficient Vision TransformerspaperFlexViT: A Flexible FPGA-based Accelerator for Edge Vision TransformerspaperQuasiMoTTo: Quasi-Monte Carlo Test-Time ScalingpaperELSAA: Efficient Low-Rank and Sparse Attention Approximation for Training TransformerspaperText Template Tokens Are Implicit Semantic Registers in Diffusion Transformerstoolup006toolsim905personjulien-cpaperKaleido: Algorithm-Hardware Co-Design for Video Diffusion Transformers by Exploiting Latent Space CorrelationspaperKroQuant: Kronecker-Structured Block Transforms for Efficient Post-Training Quantization of Diffusion TransformerstoollnewsAutoencoders: Learning Through Reconstruction - Snowflaketoolup015toolsimpaperMxGPS: Multiplex Graph Transformers for a Power Grid Foundation ModeltoolraisepersonSunMarcnewsMicrosoft unveils MAI-Image-1, its first AI model that turns words into pictures - The Eastleigh Voicetoolc901modelopenai/whisper-large-v3-turbomodeldeepseek-ai/Janus-Pro-7Btoolregister_parameterpersondependabot[bot]toolc4toolplc1802modelsentence-transformers/all-MiniLM-L6-v2paperTransformer Geometry Observatory TGO-II: Representational Similarity Observatorypersonamyerobertstoolruf013paperInvariant Learning Dynamics of Transformers in Inductive Reasoning TaskspaperMobius Learning: Cyclic Depth Folding in TransformerspersonNarsiltoolpy310personLysandreJikpersonCyrilvallezpersonpatrickvonplatentooluppaperUmm... With Transformers? Insights from Filled Pause Use across Four Slavic Parliamentstoole741personcyyevermodeldeepseek-ai/DeepSeek-R1paperMultimodal Knowledge Edit-Scoped Generalization for Online Recursive MLLM EditingtoolfpaperDo We Really Need Multimodal Emotion Language Models Larger Than 1B Parameters?modeldeepseek-ai/DeepSeek-OCRpaperANGLE: Angular Neural Generative Learning via EngressionpersonthomwolftoolexceptpersonVictorSanhpaperDeform360: A Massive Multi-view Visuotactile Dataset for Deformable World ModelspaperCAIRN: Cross-Room 3D Scene Understanding with Topology-Aware Large Multimodal ModelspersonRocketknight1personzucchini-nlpnewsFLUX 3 - Real World Models: Towards Multimodal Flow Models as the Backbone of Visual Intelligencemodeldeepseek-ai/DeepSeek-V3-0324tooltransformerspersonsanchit-gandhitoolperf102persongantepaperSVF-CR: Synchronized Visual-Facial Cross-Refinement for Multimodal Ambivalence and Hesitancy Recognitionpersonstas00personpatil-surajtoolwpersonArthurZuckertools110paperMotion4Motion: Motion Transfer Across Subjects at Inferencemodelmistralai/Mistral-7B-Instruct-v0.2modelmistralai/Mixtral-8x7B-Instruct-v0.1personstevhliunewsTransformers in Deep Learning: How Self-Attention Changed Modern AI - SnowflakepaperGrokking in small transformersmodeldeepseek-ai/DeepSeek-V4-Propersonydshiehtoolsim1personsguggertoolplc0208toolopaperPost-Training Pruning for Diffusion Transformers
