repoGitHubTrust 82 · PrimaryPublished 1mo agoLive · 1mo ago
lechmazur/writing
This benchmark tests how well LLMs incorporate a set of 10 mandatory story elements (characters, objects, core concepts, attributes, motivations, etc.) in a short creative story
Lineage graph
Paper → model → repo connections mined from source citations (Tier-1 exact match).
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.
- PossiblePossibly related (embedding) · 57%Evaluate a model properly →
- PossiblePossibly related (embedding) · 49%From Personas to Plot: Character-Grounded Multi-Agent Story Generation for Long-Form Narratives →
- PossiblePossibly related (embedding) · 46%Large Language Models Explained: How LLMs Actually Work and Who Builds the Biggest Ones - SQ Magazine →
- PossiblePossibly related (embedding) · 45%Measuring the Gap Between Human and LLM Research Ideas →
- PossiblePossibly related (embedding) · 45%Poller: Are LLMs Suitable for Evaluating the Poetry Understanding Task? →
Related to
Implements
Covers
Related across the graph
paperPoller: Are LLMs Suitable for Evaluating the Poetry Understanding Task?newsLarge Language Models Explained: How LLMs Actually Work and Who Builds the Biggest Ones - SQ MagazinepaperFrom Personas to Plot: Character-Grounded Multi-Agent Story Generation for Long-Form NarrativestutorialEvaluate a model properlypaperMeasuring the Gap Between Human and LLM Research Ideas
