repoGitHubTrust 82 · PrimaryPublished 1mo agoLive · 23d ago
openinfer-project/openinfer
Pure Rust + CUDA LLM inference engine — no PyTorch, OpenAI-compatible, serves Qwen3 to Kimi-K2
Lineage graph
Paper → model → repo connections mined from source citations (Tier-1 exact match).
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.
- PossiblePossibly related (embedding) · 54%OpenAI and Broadcom unveil LLM-optimized inference chip →
- PossiblePossibly related (embedding) · 54%OpenAI unveils its first custom chip, built by Broadcom →
- PossiblePossibly related (embedding) · 52%A barebones CPU-only inference engine for Qwen 3, written from scratch in pure C →
- PossiblePossibly related (embedding) · 48%Kuma: compiling PyTorch models into self-contained WebGPU executables [P] →
- PossiblePossibly related (embedding) · 48%OpenAI is teasing new hardware… for Codex →
- PossiblePossibly related (embedding) · 54%China's Moonshot AI claims Kimi K3 can rival OpenAI and Anthropic →
- PossiblePossibly related (embedding) · 56%543 tok/s single-request Qwen3.6-35B-A3B on one RTX 5090 over a 65K-token decode →
Covers
newsOpenAI and Broadcom unveil LLM-optimized inference chipnewsOpenAI unveils its first custom chip, built by BroadcomnewsA barebones CPU-only inference engine for Qwen 3, written from scratch in pure CnewsKuma: compiling PyTorch models into self-contained WebGPU executables [P]newsOpenAI is teasing new hardware… for Codex
Covers (incoming)
newsChina's Moonshot AI claims Kimi K3 can rival OpenAI and Anthropicnews543 tok/s single-request Qwen3.6-35B-A3B on one RTX 5090 over a 65K-token decodenewsKimi K3 in the next few hours. Deepseek V4 GA later in the week. New Liquid models. New Mistral models sometime this month. And some rumours suggest GLM 5.5 is coming in August. Openweight AI is eating good.newsMoonshot’s upcoming Kimi 3 is expected to close the gap with Anthropic’s Opus 4.8newsHigh-Performance MoE Inference: Qwen3.6–35B-A3B on an AI PC with OpenVINO - Medium
Related across the graph
newsOpenAI unveils its first custom chip, built by BroadcomnewsOpenAI and Broadcom unveil LLM-optimized inference chipnewsKuma: compiling PyTorch models into self-contained WebGPU executables [P]news543 tok/s single-request Qwen3.6-35B-A3B on one RTX 5090 over a 65K-token decodenewsOpenAI is teasing new hardware… for CodexnewsKimi K3 in the next few hours. Deepseek V4 GA later in the week. New Liquid models. New Mistral models sometime this month. And some rumours suggest GLM 5.5 is coming in August. Openweight AI is eating good.newsChina's Moonshot AI claims Kimi K3 can rival OpenAI and AnthropicnewsHigh-Performance MoE Inference: Qwen3.6–35B-A3B on an AI PC with OpenVINO - MediumnewsMoonshot’s upcoming Kimi 3 is expected to close the gap with Anthropic’s Opus 4.8newsA barebones CPU-only inference engine for Qwen 3, written from scratch in pure C
