Read original ↗
paperarXivTrust 82 · PrimaryPublished 3d agoLive · 13h ago

NL2AGBench: Benchmarking LLM Auto-Formalization for AlphaGeometry

Recent advances in large language models (LLMs) have demonstrated strong capabilities in natural language understanding and mathematical reasoning. However, their ability to translate informal mathematical problems into formal representations remains underexplored. This limitation is particularly important for neuro-symbolic geometry systems such as AlphaGeometry, whose theorem-proving engine requires inputs in a specialized domain-specific language (DSL). Although AlphaGeometry achieves near-IMO gold-medalist performance, manually converting natural-language problems into its formal syntax re

Lineage graph

Paper → model → repo connections mined from source citations (Tier-1 exact match).

Why these links exist

Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.

  • FuzzyOverlapping authors or contributors · 62%TauricResearch/TradingAgents

    Shared author/contributor keys: xiao

  • LinkedLinked via arxiv author · 85%Samuel Xiao

    NL2AGBench: Benchmarking LLM Auto-Formalization for AlphaGeometry

  • LinkedLinked via arxiv author · 85%Judy Song

    NL2AGBench: Benchmarking LLM Auto-Formalization for AlphaGeometry

  • LinkedLinked via arxiv author · 85%Rory Hu

    NL2AGBench: Benchmarking LLM Auto-Formalization for AlphaGeometry

  • LinkedLinked via arxiv author · 85%Ziliang Zong

    NL2AGBench: Benchmarking LLM Auto-Formalization for AlphaGeometry

Implements (incoming)

authored (incoming)

Related across the graph

Topics