Flesch-Kincaid Readability Depends Only on the Topic Distribution in Long Texts under Topic Models
Flesch Reading Ease (FRE) and the Flesch-Kincaid Grade Level (FKGL) are widely used readability scores for English computed from the same two document statistics, yet their stability on long documents need not imply invariance to lexical composition. Surprisingly, under a topic model with an explicit sentence-boundary token, both scores converge almost surely to deterministic functions of the document topic distribution through just two scalar rates: in the long-text limit, all score variation is mediated by topical composition rather than any residual readability signal. The theory covers bot
Lineage graph
Paper → model → repo connections mined from source citations (Tier-1 exact match).
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.
- PossiblePossibly related (embedding) · 46%Large Language Models Are Still Getting Stronger, but Researchers Face New Bottlenecks in Data, Evaluation, and Safety | Newswise - Newswise →
- FuzzySimilar title/name (fuzzy) · 87%MaartenGr/BERTopic →
“Fuzzy title match (0.94): “Flesch-Kincaid Readability Depends Only on the Topic Distrib” ≈ “MaartenGr/BERTopic””
- LinkedLinked via arxiv author · 85%Yo Ehara →
“Flesch-Kincaid Readability Depends Only on the Topic Distribution in Long Texts under Topic Models”
