Cross-lingual Biography Enrichment via Claim Extraction and Alignment
English Wikipedia is often treated as the default encyclopedic source, yet non-English Wikipedia editions can contain richer locally grounded information for long-tail figures. We study cross-lingual biography enrichment: enriching an existing English biography with facts supported by a non-English biography about the same person. Focusing on women from non-English-speaking contexts, we introduce \textsc{CLAW-4L}, a benchmark consisting of 300 Wikipedia biography pairs linking an English biography with its French, Chinese or Azerbaijani counterpart, along with claim annotations and a fine-grai
Lineage graph
Paper → model → repo connections mined from source citations (Tier-1 exact match).
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.
- PossiblePossibly related (embedding) · 46%Large Language Models Are Still Getting Stronger, but Researchers Face New Bottlenecks in Data, Evaluation, and Safety | Newswise - Newswise →
- LinkedLinked via arxiv author · 85%Yifei Song →
“Cross-lingual Biography Enrichment via Claim Extraction and Alignment”
- LinkedLinked via arxiv author · 85%Ziyang Chen →
“Cross-lingual Biography Enrichment via Claim Extraction and Alignment”
- LinkedLinked via arxiv author · 85%Emil Sayilov →
“Cross-lingual Biography Enrichment via Claim Extraction and Alignment”
- LinkedLinked via arxiv author · 85%Claire Gardent →
“Cross-lingual Biography Enrichment via Claim Extraction and Alignment”
