Skip to main content
Angestrom home
SearchPapersModelsLive AIIntelligence
SearchβŒ•βŒ˜K
EnterprisePricingSign in

Stay Ahead in the AI Revolution

Weekly digest β€” EPI pulse, top intelligence, fresh lineage. Free, no account.

Follow Angestrom
Global source network
Synced every 5 minutes

Continuous sync from primary AI sources β€” indexed, enriched, and queryable in real time.

arXivHugging FaceGitHubOpenAIAnthropicDeepMindReutersBBC TechHacker NewsReddit MLVerified feedsFunding
ANGESTROM

The Intelligence Layer of Humanity. Everything AI. All in One Place.

Angestrom connects every piece of the AI ecosystem β€” data, models, research, companies, tools, and people.

info@angestrom.comwww.angestrom.comLucknow, Uttar Pradesh, India

Product

  • AI Search
  • AI Models
  • Research Papers
  • Companies
  • News & Events
  • GitHub Explorer
  • APIs & Tools
  • Datasets
  • Benchmarks
  • Model lifecycle
  • Funding graph
  • Contributors
  • AI Agents

Resources

  • Weekly digest
  • Documentation
  • Tutorials
  • Guides
  • News
  • Help / Start
  • Community

Company

  • About
  • Contact
  • Privacy Policy
  • Terms of Service
  • Acceptable Use

Enterprise

  • Pricing
  • Workspace
  • Contact Sales

Developer

  • Developer Hub
  • API docs
  • GitHub

Learn

  • Learning Academy
  • Roadmaps
  • Glossary
  • AI for Beginners

Popular Topics

Loading topics…
View All Topics β†’
Β© 2026 Angestrom Intelligence Private Limited. All rights reserved.
English
Theme
Angestrom home
SearchPapersModelsLive AIIntelligence
SearchβŒ•βŒ˜K
EnterprisePricingSign in
  1. Home
  2. /Repositories
  3. /huggingface/transformers
Read original β†—
repoGitHubTrust 82 Β· PrimaryPublished 1mo agoLive Β· 10h ago

huggingface/transformers

πŸ€— Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and training.

Lineage graph

Paper β†’ model β†’ repo connections mined from source citations (Tier-1 exact match).

Why these links exist

Every edge carries a method, confidence, and the source snippet that justified it β€” so bad links are debuggable.

  • LinkedLinked via dependency parse Β· 85%c4 β†’

    β€œListed in pyproject.toml”

  • LinkedLinked via dependency parse Β· 85%c901 β†’

    β€œListed in pyproject.toml”

  • LinkedLinked via dependency parse Β· 85%e β†’

    β€œListed in pyproject.toml”

  • LinkedLinked via dependency parse Β· 85%e501 β†’

    β€œListed in pyproject.toml”

  • LinkedLinked via dependency parse Β· 85%e741 β†’

    β€œListed in pyproject.toml”

  • LinkedLinked via dependency parse Β· 85%except β†’

    β€œListed in pyproject.toml”

  • LinkedLinked via dependency parse Β· 85%f β†’

    β€œListed in pyproject.toml”

  • LinkedLinked via dependency parse Β· 85%furb β†’

    β€œListed in pyproject.toml”

Related to

toolc4toolc901tooletoole501toole741toolexcepttoolftoolfurbtoolitoolltoolotoolperf102toolpie794toolplc0208toolplc1802toolpy310toolraisetoolregister_parametertoolruf013tools110toolsimtoolsim1toolsim905tooltransformerstooluptoolup006toolup015toolup031toolw

Implements

paperGrokking in small transformerspaperInhibited Self-Attention: Sharpening Focus in Vision TransformerspaperAutomated Compliance Mapping in Cloud Security with Domain-Adapted Sentence TransformerspaperPAC-ACT: Post-training Actor-Critic for Action Chunking TransformerspaperFoveation-Guided Dynamic Token Selection for Robust and Efficient Vision TransformerspaperFrom Expressivity to Sample Complexity: Narrow Teachers for Transformers via C-RASPpaperQuasiMoTTo: Quasi-Monte Carlo Test-Time ScalingpaperInvariant Learning Dynamics of Transformers in Inductive Reasoning TaskspaperPost-Training Pruning for Diffusion TransformerspaperMobius Learning: Cyclic Depth Folding in TransformerspaperAppearance Pointers -- Multimodal Region Control of Diffusion TransformerspaperText Template Tokens Are Implicit Semantic Registers in Diffusion TransformerspaperELSAA: Efficient Low-Rank and Sparse Attention Approximation for Training TransformerspaperKaleido: Algorithm-Hardware Co-Design for Video Diffusion Transformers by Exploiting Latent Space CorrelationspaperMxGPS: Multiplex Graph Transformers for a Power Grid Foundation ModelpaperFlexViT: A Flexible FPGA-based Accelerator for Edge Vision TransformerspaperReview Residuals: Update-Conditioned Residual Gating for TransformerspaperUmm... With Transformers? Insights from Filled Pause Use across Four Slavic ParliamentspaperKroQuant: Kronecker-Structured Block Transforms for Efficient Post-Training Quantization of Diffusion Transformers

Related to (incoming)

modelmistralai/Mixtral-8x7B-Instruct-v0.1modelmicrosoft/phi-2modelzai-org/GLM-5.2modelmistralai/Mistral-7B-Instruct-v0.2modelopenai/whisper-large-v3-turbomodeldeepseek-ai/DeepSeek-V3modeldeepseek-ai/DeepSeek-R1modeldeepseek-ai/DeepSeek-V4-Promodeldeepseek-ai/Janus-Pro-7BpaperA Large-Scale Measurement of AI Bill of Materials Completeness in Hugging Face Modelsmodelsentence-transformers/all-MiniLM-L6-v2modeldeepseek-ai/DeepSeek-V3-0324modelBAAI/bge-m3modeldeepseek-ai/DeepSeek-OCR

Implements (incoming)

paperMultimodal Knowledge Edit-Scoped Generalization for Online Recursive MLLM EditingpaperTransformer Geometry Observatory TGO-II: Representational Similarity ObservatorypaperDeform360: A Massive Multi-view Visuotactile Dataset for Deformable World ModelspaperCAIRN: Cross-Room 3D Scene Understanding with Topology-Aware Large Multimodal ModelspaperMotion4Motion: Motion Transfer Across Subjects at InferencepaperANGLE: Angular Neural Generative Learning via EngressionpaperDo We Really Need Multimodal Emotion Language Models Larger Than 1B Parameters?paperXRFormer: Multiscale Tokenization for XRF Representation LearningpaperSVF-CR: Synchronized Visual-Facial Cross-Refinement for Multimodal Ambivalence and Hesitancy Recognition

contributed_to (incoming)

personydshiehpersonthomwolfpersonsguggerpersonLysandreJikpersonpatrickvonplatenpersongantepersonstas00personArthurZuckerpersonzucchini-nlppersonRocketknight1personjulien-cpersonCyrilvallezpersonyounesbelkadapersonNielsRoggepersonsshleiferpersonSunMarcpersonamyerobertspersonNarsilpersonstevhliupersonpatil-surajpersonVictorSanhpersoncyyeverpersonyonigozlanpersondependabot[bot]personsanchit-gandhi

Covers (incoming)

newsTransformers in Deep Learning: How Self-Attention Changed Modern AI - SnowflakenewsMicrosoft unveils MAI-Image-1, its first AI model that turns words into pictures - The Eastleigh VoicenewsAutoencoders: Learning Through Reconstruction - SnowflakenewsFLUX 3 - Real World Models: Towards Multimodal Flow Models as the Backbone of Visual Intelligence

Related across the graph

toolipersonyounesbelkadapersonsshleifertoolup031paperReview Residuals: Update-Conditioned Residual Gating for TransformerspaperAutomated Compliance Mapping in Cloud Security with Domain-Adapted Sentence TransformerstoolfurbpaperInhibited Self-Attention: Sharpening Focus in Vision TransformersmodelBAAI/bge-m3toolepersonNielsRoggepaperA Large-Scale Measurement of AI Bill of Materials Completeness in Hugging Face Modelspersonyonigozlantoole501modelzai-org/GLM-5.2modelmicrosoft/phi-2paperAppearance Pointers -- Multimodal Region Control of Diffusion TransformerspaperXRFormer: Multiscale Tokenization for XRF Representation LearningpaperPAC-ACT: Post-training Actor-Critic for Action Chunking Transformersmodeldeepseek-ai/DeepSeek-V3toolpie794paperFrom Expressivity to Sample Complexity: Narrow Teachers for Transformers via C-RASPpaperFoveation-Guided Dynamic Token Selection for Robust and Efficient Vision TransformerspaperFlexViT: A Flexible FPGA-based Accelerator for Edge Vision TransformerspaperQuasiMoTTo: Quasi-Monte Carlo Test-Time ScalingpaperELSAA: Efficient Low-Rank and Sparse Attention Approximation for Training TransformerspaperText Template Tokens Are Implicit Semantic Registers in Diffusion Transformerstoolup006toolsim905personjulien-cpaperKaleido: Algorithm-Hardware Co-Design for Video Diffusion Transformers by Exploiting Latent Space CorrelationspaperKroQuant: Kronecker-Structured Block Transforms for Efficient Post-Training Quantization of Diffusion TransformerstoollnewsAutoencoders: Learning Through Reconstruction - Snowflaketoolup015toolsimpaperMxGPS: Multiplex Graph Transformers for a Power Grid Foundation ModeltoolraisepersonSunMarcnewsMicrosoft unveils MAI-Image-1, its first AI model that turns words into pictures - The Eastleigh Voicetoolc901modelopenai/whisper-large-v3-turbomodeldeepseek-ai/Janus-Pro-7Btoolregister_parameterpersondependabot[bot]toolc4toolplc1802modelsentence-transformers/all-MiniLM-L6-v2paperTransformer Geometry Observatory TGO-II: Representational Similarity Observatorypersonamyerobertstoolruf013paperInvariant Learning Dynamics of Transformers in Inductive Reasoning TaskspaperMobius Learning: Cyclic Depth Folding in TransformerspersonNarsiltoolpy310personLysandreJikpersonCyrilvallezpersonpatrickvonplatentooluppaperUmm... With Transformers? Insights from Filled Pause Use across Four Slavic Parliamentstoole741personcyyevermodeldeepseek-ai/DeepSeek-R1paperMultimodal Knowledge Edit-Scoped Generalization for Online Recursive MLLM EditingtoolfpaperDo We Really Need Multimodal Emotion Language Models Larger Than 1B Parameters?modeldeepseek-ai/DeepSeek-OCRpaperANGLE: Angular Neural Generative Learning via EngressionpersonthomwolftoolexceptpersonVictorSanhpaperDeform360: A Massive Multi-view Visuotactile Dataset for Deformable World ModelspaperCAIRN: Cross-Room 3D Scene Understanding with Topology-Aware Large Multimodal ModelspersonRocketknight1personzucchini-nlpnewsFLUX 3 - Real World Models: Towards Multimodal Flow Models as the Backbone of Visual Intelligencemodeldeepseek-ai/DeepSeek-V3-0324tooltransformerspersonsanchit-gandhitoolperf102persongantepaperSVF-CR: Synchronized Visual-Facial Cross-Refinement for Multimodal Ambivalence and Hesitancy Recognitionpersonstas00personpatil-surajtoolwpersonArthurZuckertools110paperMotion4Motion: Motion Transfer Across Subjects at Inferencemodelmistralai/Mistral-7B-Instruct-v0.2modelmistralai/Mixtral-8x7B-Instruct-v0.1personstevhliunewsTransformers in Deep Learning: How Self-Attention Changed Modern AI - SnowflakepaperGrokking in small transformersmodeldeepseek-ai/DeepSeek-V4-Propersonydshiehtoolsim1personsguggertoolplc0208toolopaperPost-Training Pruning for Diffusion Transformers
Knowledge path··i→·younesbelkada→·sshleifer→Rhuggingface/transformers

Topics

audiodeep-learningdeepseekgemmaglmhacktoberfestllmmachine-learningmodel-hubnatural-language-processing

Explore

Search similar β†’Knowledge graph β†’All repos β†’Full intelligence feed β†’
Maintenance98
RIS100GitHub verified
Graph trust82Primary
Graph score164231