repoGitHubTrust 82 · PrimaryPublished 12d agoLive · 8d ago
agentscope-ai/OpenJudge
OpenJudge: A Unified Framework for Holistic Evaluation and Quality Rewards
Lineage graph
Paper → model → repo connections mined from source citations (Tier-1 exact match).
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.
- PossiblePossibly related (embedding) · 62%Research acceleration: The view inside OpenAI →
- PossiblePossibly related (embedding) · 60%Helping build shared standards for advanced AI →
- PossiblePossibly related (embedding) · 58%The agent evaluation gap: Enterprise AI organizations have a reality-alignment problem, not a coverage problem — and most are shipping to production anyway →
- PossiblePossibly related (embedding) · 58%Beyond benchmarks: The 5 pillars of AI evaluation systems →
- PossiblePossibly related (embedding) · 63%New Open Source Benchmark Scores AI Agents on Their Ability to Learn and Perform Complex Actions - WBOC TV →
Covers
newsResearch acceleration: The view inside OpenAInewsHelping build shared standards for advanced AInewsThe agent evaluation gap: Enterprise AI organizations have a reality-alignment problem, not a coverage problem — and most are shipping to production anywaynewsBeyond benchmarks: The 5 pillars of AI evaluation systems
Covers (incoming)
Related across the graph
newsResearch acceleration: The view inside OpenAInewsNew Open Source Benchmark Scores AI Agents on Their Ability to Learn and Perform Complex Actions - WBOC TVnewsThe agent evaluation gap: Enterprise AI organizations have a reality-alignment problem, not a coverage problem — and most are shipping to production anywaynewsBeyond benchmarks: The 5 pillars of AI evaluation systemsnewsHelping build shared standards for advanced AI
