repoGitHubTrust 82 · PrimaryPublished 1mo agoLive · 24d ago
modelscope/evalscope
A streamlined and customizable framework for efficient large model (LLM, VLM, AIGC) evaluation and performance benchmarking.
Lineage graph
Paper → model → repo connections mined from source citations (Tier-1 exact match).
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.
- PossiblePossibly related (embedding) · 61%Evaluate a model properly →
- PossiblePossibly related (embedding) · 57%EvalBoard →
- PossiblePossibly related (embedding) · 51%Is it agentic enough? Benchmarking open models on your own tooling →
- PossiblePossibly related (embedding) · 48%PACE: A Proxy for Agentic Capability Evaluation →
- PossiblePossibly related (embedding) · 48%Helix-7B →
- PossiblePossibly related (embedding) · 63%Best Local VLMs - July 2026 →
