newsReddit r/LocalLLaMATrust 52 · CommunityPublished 1mo agoLive · 1mo ago
Are there good closed vs open LLM rankings? Also, are 70B–350B models actually worth it?
hey, I’m currently getting enough VRAM to run something in the GLM-5.2 range, but I’m wondering: do we actually have a solid ranking that compares closed-source and open-weight LLMs side by side? I’ve been trying to find a clear “closed vs open” leaderboard, but most benchmarks feel fragmented or don’t really answer the practical question of what’s actually best to run locally versus what’s only competitive through API models. Also, are the
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.
- LinkedLinked via unknownyyh-001/llm-value-rankings →
- LinkedLinked via unknownEvaluate a model properly →
- LinkedLinked via unknownSurrogate Fidelity: When Can Open LLMs Explain Closed Ones? →
- PossiblePossibly related (embedding) · 45%qualcomm/GenieX →
- PossiblePossibly related (embedding) · 50%open-compass/opencompass →
- PossiblePossibly related (embedding) · 50%notwitcheer/llm-bench-rig →
- PossiblePossibly related (embedding) · 46%BodhiSearch/BodhiApp →
- PossiblePossibly related (embedding) · 51%development-and-operations/model-fit →
Covers
Covers (incoming)
Related across the graph
repogqgs/llm100kbenchrepollm-ring/lmringreponotwitcheer/llm-bench-rigrepoopen-compass/opencompassrepodevelopment-and-operations/model-fitpaperSurrogate Fidelity: When Can Open LLMs Explain Closed Ones?repoManasVardhan/bench-my-llmtutorialEvaluate a model properlyrepoBodhiSearch/BodhiApprepoyyh-001/llm-value-rankingsrepoqualcomm/GenieX
