Verifiable Self-Evolution for Open-Ended Dialogue Skills via Future-Feedback Prediction
Textual skills provide a lightweight way to improve frozen language-model agents, but their self-evolution normally requires a stable validation signal. Such signals are natural in mathematics or code, where an answer can be checked after it changes, yet are problematic in open-ended dialogue: changing the assistant response also changes the user's next reaction, so a logged reaction cannot directly evaluate a counterfactual response. We propose future-feedback skill evolution, which first redirects self-evolution from prescribing the current answer to predicting whether the observed answer wi
Lineage graph
Paper → model → repo connections mined from source citations (Tier-1 exact match).
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.
- PossiblePossibly related (embedding) · 50%Superhuman’s new auto-draft feature almost makes me like AI replies →
- FuzzyOverlapping authors or contributors · 62%affaan-m/ECC →
“Shared author/contributor keys: jiang”
- FuzzyOverlapping authors or contributors · 62%BerriAI/litellm →
“Shared author/contributor keys: jiang”
- FuzzySimilar title/name (fuzzy) · 59%KKKKhazix/khazix-skills →
“Fuzzy title match (0.73): “Verifiable Self-Evolution for Open-Ended Dialogue Skills via” ≈ “KKKKhazix/khazix-skills””
- LinkedLinked via arxiv author · 85%ChaoJin Zhao →
“Verifiable Self-Evolution for Open-Ended Dialogue Skills via Future-Feedback Prediction”
- LinkedLinked via arxiv author · 85%Xuan Jiang →
“Verifiable Self-Evolution for Open-Ended Dialogue Skills via Future-Feedback Prediction”
