repoGitHubTrust 82 · PrimaryPublished 3d agoLive · 3d ago
mlc-ai/web-llm
High-performance In-browser LLM Inference Engine
Lineage graph
Paper → model → repo connections mined from source citations (Tier-1 exact match).
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.
- PossiblePossibly related (embedding) · 83%WebLLM: high-performance in-browser LLM inference engine →
- PossiblePossibly related (embedding) · 61%OpenAI and Broadcom unveil LLM-optimized inference chip →
- PossiblePossibly related (embedding) · 59%OpenAI and Broadcom announce chip designed for LLM inference at scale →
- PossiblePossibly related (embedding) · 56%LLMs could control their host machines by exploiting inference engines →
- PossiblePossibly related (embedding) · 58%Benchmarking small LLM inference on SageMaker AI: G7 vs G5 and G6 - aws.amazon.com →
- PossiblePossibly related (embedding) · 31%Pre-Compiled Pipeline Shards for Distributed LLM Inference on Intel AI PC Fleets →
“Possibly related via embedding similarity 0.60 (not asserted). Timestamp check: artifact after paper (+19d).”
- PossiblePossibly related (embedding) · 62%Benchmarking small LLM inference on SageMaker AI: G7 vs G5 and G6 - Amazon Web Services (AWS) →
Covers
Covers (incoming)
Related to (incoming)
Related across the graph
newsBenchmarking small LLM inference on SageMaker AI: G7 vs G5 and G6 - aws.amazon.comnewsOpenAI and Broadcom announce chip designed for LLM inference at scalenewsOpenAI and Broadcom unveil LLM-optimized inference chipnewsLLMs could control their host machines by exploiting inference enginesnewsWebLLM: high-performance in-browser LLM inference enginenewsBenchmarking small LLM inference on SageMaker AI: G7 vs G5 and G6 - Amazon Web Services (AWS)paperPre-Compiled Pipeline Shards for Distributed LLM Inference on Intel AI PC Fleets
