repoGitHubTrust 82 · PrimaryPublished 1mo agoLive · yesterday
vllm-project/vllm
A high-throughput and memory-efficient inference and serving engine for LLMs
Lineage graph
Paper → model → repo connections mined from source citations (Tier-1 exact match).
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.
- PossiblePossibly related (embedding) · 60%OpenAI and Broadcom announce chip designed for LLM inference at scale →
- PossiblePossibly related (embedding) · 56%OpenAI and Broadcom unveil LLM-optimized inference chip →
- PossiblePossibly related (embedding) · 55%Hardware startup unveils inference accelerator →
- PossiblePossibly related (embedding) · 54%Evaluate a model properly →
- PossiblePossibly related (embedding) · 50%IEEE Rolls Out Large Language Models Virtual Training Course →
- LinkedLinked via dependency parse · 85%all →
“Listed in pyproject.toml”
- LinkedLinked via dependency parse · 85%apache-2.0 →
“Listed in pyproject.toml”
- LinkedLinked via dependency parse · 85%b →
“Listed in pyproject.toml”
Covers
Related to
tutorialEvaluate a model properlytoolalltoolapache-2.0toolbtoolb007toolb905tooldependenciestooletoole731toolftoolf401toolf403toolf405toolgtoolitoolisctooljinja2toollicensetoolninjatooloptional-dependenciestoolpydantic.mypytoolreadme.mdtoolsetuptools.build_metatoolsimtooluptoolup032toolversiontoolvllmtoolvllm.general_pluginstoolwheel
Related to (incoming)
modeldeepseek-ai/DeepSeek-R1modelmistralai/Mixtral-8x7B-Instruct-v0.1modeldeepseek-ai/DeepSeek-V3modelzai-org/GLM-5.2paperFreqDepthKV: Frequency-Guided Depth Sharing for Robust KV Cache Compression in Long-Context LLM Inferencemodelbaidu/Unlimited-OCRmodelmoonshotai/Kimi-K3modelQwen/Qwen3.8-27BmodelMiniMaxAI/MiniMax-H3
Implements (incoming)
contributed_to (incoming)
personDarkLight1337personWoosukKwonpersonmgoinpersonhmellorpersonyoukaichaopersonIsotr0pypersonnjhillpersonyewentao256personjeejeeleepersonAndreasKaratzaspersonywang96personLucasWilkinsonpersonNickLucchepersonsimon-mopersonrussellbpersonrobertgshaw2-redhatpersonchaunceyjiangpersonnooooppersonreidliu41personkhluupersontlrmchlsmthpersonzhuohan123personMatthewBonannipersonbigPYJ1151personjikunshang
Covers (incoming)
Related across the graph
toolimodelQwen/Qwen3.8-27Btoolvllmpersonkhluupersonnjhillpersonywang96personrussellbmodelmoonshotai/Kimi-K3toolwheeltooletooldependenciesmodelzai-org/GLM-5.2newsOpenAI and Broadcom announce chip designed for LLM inference at scalenewsCloud-vLLM Benchmark Differences [R]paperAn Efficient vLLM-Based Inference Pipeline for Unified Audio Understanding and GenerationpersonDarkLight1337toollicensepersonIsotr0pymodeldeepseek-ai/DeepSeek-V3modelMiniMaxAI/MiniMax-H3personsimon-monewsOpenAI and Broadcom unveil LLM-optimized inference chippersonbigPYJ1151toolup032personyewentao256toole731toolvllm.general_pluginspersonjikunshangtoolgpersonLucasWilkinsontoolsimtooloptional-dependenciespersonrobertgshaw2-redhatnewsCan LLMs Perform Deep Technical Comprehension of Computer Architecture PaperstoolninjanewsTried testing qwen 35b moe model on s26 ultra , without compromising on precision [R] ,[D]toolpydantic.mypypersonchaunceyjiangpaperFreqDepthKV: Frequency-Guided Depth Sharing for Robust KV Cache Compression in Long-Context LLM Inferencetoolversiontoolapache-2.0persontlrmchlsmthtoolbtooluppersonAndreasKaratzastoolalltoolf405modeldeepseek-ai/DeepSeek-R1toolftooljinja2personmgointoolreadme.mdpersonNickLucchetoolf401newsIEEE Rolls Out Large Language Models Virtual Training Coursepersonzhuohan123personjeejeeleetoolisctutorialEvaluate a model properlynewsHardware startup unveils inference acceleratorpersonWoosukKwonpersonMatthewBonannimodelbaidu/Unlimited-OCRtoolf403modelmistralai/Mixtral-8x7B-Instruct-v0.1personnooooptoolb007toolsetuptools.build_metapersonyoukaichaopersonreidliu41toolb905personhmellor
