repoGitHubTrust 82 · PrimaryPublished 1mo agoLive · 1mo ago
AgentToolkit/altk-evolve
Self improving agents through iterations
Lineage graph
Paper → model → repo connections mined from source citations (Tier-1 exact match).
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.
- PossiblePossibly related (embedding) · 62%I made a superhuman Generals.io agent with self-play RL [P] →
- PossiblePossibly related (embedding) · 60%EvoPolicyGym: Evaluating Autonomous Policy Evolution in Interactive Environments →
- PossiblePossibly related (embedding) · 59%Self-rewarding agents that retrace failures →
- PossiblePossibly related (embedding) · 56%Learning from Failure: Inference-Time Self-Improvement for Computer-Use Agents →
- PossiblePossibly related (embedding) · 56%Alibaba's model never trained as an agent — and improved agent performance across seven benchmarks →
- PossiblePossibly related (embedding) · 51%EdgeBench: Unveiling Scaling Laws of Learning from Real-World Environments →
- PossiblePossibly related (embedding) · 64%MetaSkill-Evolve: Recursive Self-Improvement of LLM Agents via Two-Timescale Meta-Skill Evolution →
- PossiblePossibly related (embedding) · 46%Alex Krentsel on Exo and Recursive Self Improving A… - StartupHub.ai →
Covers
Implements
Implements (incoming)
paperEdgeBench: Unveiling Scaling Laws of Learning from Real-World EnvironmentspaperMetaSkill-Evolve: Recursive Self-Improvement of LLM Agents via Two-Timescale Meta-Skill EvolutionpaperCurateEvo: Data-Curation Evolving for Agentic Post-TrainingpaperWho Grades the Grader? Co-Evolving Evaluation Metrics and Skills for Self-Improving LLM Agents
Covers (incoming)
Related across the graph
news[2607.07508] Single-Rollout Asynchronous Optimization for Agentic Reinforcement LearningpaperCurateEvo: Data-Curation Evolving for Agentic Post-TrainingpaperSelf-rewarding agents that retrace failurespaperMetaSkill-Evolve: Recursive Self-Improvement of LLM Agents via Two-Timescale Meta-Skill EvolutionpaperWho Grades the Grader? Co-Evolving Evaluation Metrics and Skills for Self-Improving LLM AgentspaperEdgeBench: Unveiling Scaling Laws of Learning from Real-World EnvironmentsnewsHow self-improving harnesses are rewriting the agent engineering playbook - TechTalkspaperLearning from Failure: Inference-Time Self-Improvement for Computer-Use AgentsnewsI made a superhuman Generals.io agent with self-play RL [P]newsAlibaba's model never trained as an agent — and improved agent performance across seven benchmarkspaperEvoPolicyGym: Evaluating Autonomous Policy Evolution in Interactive EnvironmentsnewsAlex Krentsel on Exo and Recursive Self Improving A… - StartupHub.ai
