repoGitHubTrust 82 Β· PrimaryPublished 25d agoLive Β· 23d ago
huggingface/optimum-intel
π€ Optimum Intel: Accelerate inference with Intel optimization tools
Lineage graph
Paper β model β repo connections mined from source citations (Tier-1 exact match).
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it β so bad links are debuggable.
- PossiblePossibly related (embedding) Β· 59%Hardware startup unveils inference accelerator β
- PossiblePossibly related (embedding) Β· 54%[Paper] Automated Tensor Scheduling for Hybrid CPU-GPU LLM Inference on Consumer Devices β
- PossiblePossibly related (embedding) Β· 53%DeepSeek open-sources inference optimizations with 60β85% faster generation [pdf] β
- PossiblePossibly related (embedding) Β· 52%OpenAI and Broadcom unveil LLM-optimized inference chip β
- PossiblePossibly related (embedding) Β· 52%Intel and Google deepen AI ties for chip design - thestreet.com β
- PossiblePossibly related (embedding) Β· 49%SkewAdam: A tiered optimizer that cuts MoE state memory by 97% (fits a 6.7B MoE on a 40GB GPU) [R] β
- PossiblePossibly related (embedding) Β· 49%CPU-only inference on a Celeron N5095 SBC: 6 models from 0.6B to 8B, benchmarked β
Covers
newsHardware startup unveils inference acceleratornews[Paper] Automated Tensor Scheduling for Hybrid CPU-GPU LLM Inference on Consumer DevicesnewsDeepSeek open-sources inference optimizations with 60β85% faster generation [pdf]newsOpenAI and Broadcom unveil LLM-optimized inference chipnewsIntel and Google deepen AI ties for chip design - thestreet.com
Covers (incoming)
Related across the graph
newsOpenAI and Broadcom unveil LLM-optimized inference chipnews[Paper] Automated Tensor Scheduling for Hybrid CPU-GPU LLM Inference on Consumer DevicesnewsIntel and Google deepen AI ties for chip design - thestreet.comnewsCPU-only inference on a Celeron N5095 SBC: 6 models from 0.6B to 8B, benchmarkednewsSkewAdam: A tiered optimizer that cuts MoE state memory by 97% (fits a 6.7B MoE on a 40GB GPU) [R]newsDeepSeek open-sources inference optimizations with 60β85% faster generation [pdf]newsHardware startup unveils inference accelerator
