Skip to main content
Angestrom home
SearchPapersModelsLive AIIntelligence
Search⌕⌘K
EnterprisePricingSign in

Stay Ahead in the AI Revolution

Weekly digest — EPI pulse, top intelligence, fresh lineage. Free, no account.

Follow Angestrom
Global source network
Synced every 5 minutes

Continuous sync from primary AI sources — indexed, enriched, and queryable in real time.

arXivHugging FaceGitHubOpenAIAnthropicDeepMindReutersBBC TechHacker NewsReddit MLVerified feedsFunding
Angestrom

Angestrom connects every piece of the AI ecosystem — data, models, research, companies, tools, and people.

info@angestrom.comwww.angestrom.comLucknow, Uttar Pradesh, India

Product

  • AI Search
  • AI Models
  • Research Papers
  • Companies
  • News & Events
  • GitHub Explorer
  • APIs & Tools
  • Datasets
  • Benchmarks
  • Model lifecycle
  • Funding graph
  • Contributors
  • AI Agents

Resources

  • Weekly digest
  • Documentation
  • Tutorials
  • Guides
  • News
  • Help / Start
  • Community

Company

  • About
  • Contact
  • Privacy Policy
  • Terms of Service
  • Acceptable Use

Enterprise

  • Pricing
  • Workspace
  • Contact Sales

Developer

  • Developer Hub
  • API docs
  • GitHub

Learn

  • Learning Academy
  • Roadmaps
  • Glossary
  • AI for Beginners

Popular Topics

Loading topics…
View All Topics →
© 2026 Angestrom. All rights reserved.
English
Theme
Angestrom home
SearchPapersModelsLive AIIntelligence
Search⌕⌘K
EnterprisePricingSign in
  1. Home
  2. /Repositories
  3. /artalis-io/bitnet.c
Read original ↗
repoGitHubTrust 82 · PrimaryPublished 1mo agoLive · 1mo ago

artalis-io/bitnet.c

Minimal, zero-dependency LLM inference in pure C11. CPU-first with NEON/AVX2 SIMD. Flash MoE (pread + LRU expert cache). TurboQuant 3-bit KV compression (8.9x less memory per session). 20+ GGUF quant formats. Compiles to WASM.

Lineage graph

Paper → model → repo connections mined from source citations (Tier-1 exact match).

Why these links exist

Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.

  • PossiblePossibly related (embedding) · 60%A barebones CPU-only inference engine for Qwen 3, written from scratch in pure C →
  • PossiblePossibly related (embedding) · 51%OpenAI and Broadcom announce chip designed for LLM inference at scale →
  • PossiblePossibly related (embedding) · 50%BiSCo-LLM: Lookup-Free Binary Spherical Coding for Extreme Low-Bit Large Language Model Compression →
  • PossiblePossibly related (embedding) · 51%Recent llama.cpp updates for SYCL/Intel →
  • PossiblePossibly related (embedding) · 47%Show HN: misa77 - a codec that decodes 2x faster than LZ4 (at better ratios) →
  • PossiblePossibly related (embedding) · 50%DeepSeek V4 Flash | IQ3_XXS-AS & IQ2_S Bench | mainline b10064 vs fairydreaming | 1xRTX 3090 + 128GB DDR4 | 250PP/11TG on 50K CTX →

Covers

newsA barebones CPU-only inference engine for Qwen 3, written from scratch in pure CnewsOpenAI and Broadcom announce chip designed for LLM inference at scale

Implements

paperBiSCo-LLM: Lookup-Free Binary Spherical Coding for Extreme Low-Bit Large Language Model Compression

Covers (incoming)

newsRecent llama.cpp updates for SYCL/IntelnewsShow HN: misa77 - a codec that decodes 2x faster than LZ4 (at better ratios)newsDeepSeek V4 Flash | IQ3_XXS-AS & IQ2_S Bench | mainline b10064 vs fairydreaming | 1xRTX 3090 + 128GB DDR4 | 250PP/11TG on 50K CTX

Related across the graph

paperBiSCo-LLM: Lookup-Free Binary Spherical Coding for Extreme Low-Bit Large Language Model CompressionnewsOpenAI and Broadcom announce chip designed for LLM inference at scalenewsRecent llama.cpp updates for SYCL/IntelnewsShow HN: misa77 - a codec that decodes 2x faster than LZ4 (at better ratios)newsA barebones CPU-only inference engine for Qwen 3, written from scratch in pure CnewsDeepSeek V4 Flash | IQ3_XXS-AS & IQ2_S Bench | mainline b10064 vs fairydreaming | 1xRTX 3090 + 128GB DDR4 | 250PP/11TG on 50K CTX
Knowledge path·PBiSCo-LLM: Lookup-Free Binary Spherical Coding for Extreme Low-Bit Large Language Model Compression→NOpenAI and Broadcom announce chip designed for LLM inference at scale→NRecent llama.cpp updates for SYCL/Intel→Rartalis-io/bitnet.c

Topics

avx2ccpu-inferenceggufinferencekv-cachellmmoeneonquantization

Explore

Search similar →Knowledge graph →All repos →Full intelligence feed →
Graph trust82Primary
Graph score22