Skip to main content
Angestrom home
SearchPapersModelsLive AIIntelligence
Search⌕⌘K
EnterprisePricingSign in

Stay Ahead in the AI Revolution

Weekly digest — EPI pulse, top intelligence, fresh lineage. Free, no account.

Follow Angestrom
Global source network
Synced every 5 minutes

Continuous sync from primary AI sources — indexed, enriched, and queryable in real time.

arXivHugging FaceGitHubOpenAIAnthropicDeepMindReutersBBC TechHacker NewsReddit MLVerified feedsFunding
ANGESTROM

The Intelligence Layer of Humanity. Everything AI. All in One Place.

Angestrom connects every piece of the AI ecosystem — data, models, research, companies, tools, and people.

info@angestrom.comwww.angestrom.comLucknow, Uttar Pradesh, India

Product

  • AI Search
  • AI Models
  • Research Papers
  • Companies
  • News & Events
  • GitHub Explorer
  • APIs & Tools
  • Datasets
  • Benchmarks
  • Model lifecycle
  • Funding graph
  • Contributors
  • AI Agents

Resources

  • Weekly digest
  • Documentation
  • Tutorials
  • Guides
  • News
  • Help / Start
  • Community

Company

  • About
  • Contact
  • Privacy Policy
  • Terms of Service
  • Acceptable Use

Enterprise

  • Pricing
  • Workspace
  • Contact Sales

Developer

  • Developer Hub
  • API docs
  • GitHub

Learn

  • Learning Academy
  • Roadmaps
  • Glossary
  • AI for Beginners

Popular Topics

Loading topics…
View All Topics →
© 2026 Angestrom Intelligence Private Limited. All rights reserved.
English
Theme
Angestrom home
SearchPapersModelsLive AIIntelligence
Search⌕⌘K
EnterprisePricingSign in
  1. Home
  2. /Models
  3. /AgentCore-8B
Read original ↗
modelHugging FaceTrust 88 · LabPublished 3mo agoLive · 3mo ago1 graph score

AgentCore-8B

A model post-trained for tool use and multi-step planning.

agents

Lineage graph

Paper → model → repo connections mined from source citations (Tier-1 exact match).

Why these links exist

Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.

  • FuzzySimilar title/name (fuzzy) · 87%zhayujie/CowAgent →

    “Fuzzy title match (0.94): “AgentCore-8B” ≈ “zhayujie/CowAgent””

  • FuzzySimilar title/name (fuzzy) · 59%NousResearch/hermes-agent →

    “Fuzzy title match (0.73): “AgentCore-8B” ≈ “NousResearch/hermes-agent””

  • FuzzySimilar title/name (fuzzy) · 59%open-multi-agent/open-multi-agent →

    “Fuzzy title match (0.73): “AgentCore-8B” ≈ “open-multi-agent/open-multi-agent””

  • FuzzySimilar title/name (fuzzy) · 59%2FastLabs/agent-squad →

    “Fuzzy title match (0.73): “AgentCore-8B” ≈ “2FastLabs/agent-squad””

  • FuzzySimilar title/name (fuzzy) · 59%TencentCloud/TencentDB-Agent-Memory →

    “Fuzzy title match (0.73): “AgentCore-8B” ≈ “TencentCloud/TencentDB-Agent-Memory””

  • FuzzySimilar title/name (fuzzy) · 87%SWE-agent/SWE-agent →

    “Fuzzy title match (0.94): “AgentCore-8B” ≈ “SWE-agent/SWE-agent””

  • FuzzySimilar title/name (fuzzy) · 59%iflytek/astron-agent →

    “Fuzzy title match (0.73): “AgentCore-8B” ≈ “iflytek/astron-agent””

  • FuzzySimilar title/name (fuzzy) · 59%bojieli/ai-agent-book →

    “Fuzzy title match (0.73): “AgentCore-8B” ≈ “bojieli/ai-agent-book””

Related to

repozhayujie/CowAgentrepoNousResearch/hermes-agentrepoopen-multi-agent/open-multi-agentrepo2FastLabs/agent-squadrepoTencentCloud/TencentDB-Agent-MemoryrepoSWE-agent/SWE-agentrepoiflytek/astron-agentrepobojieli/ai-agent-bookrepocamel-ai/owl

Covers (incoming)

newsAlibaba's model never trained as an agent — and improved agent performance across seven benchmarksnewsAgentic AI for Robot TeamsnewsOptimising LMAPF guidance graphs using Evolutionary algorithms: Advice needed [R]newsAI coding agents taught robots how to install GPUs and cut zip tiesnewsLearning to lead in a hybrid human-AI enterprisenewsNVIDIA Brings Trusted, 24/7 AI Agents to Telecom OperationsnewsAmazon Bedrock AgentCore harness is now generally available: Go from idea to production-grade agent in minutesnewsIs it agentic enough? Benchmarking open models on your own toolingnewsOpen-source agent framework crosses 50k starsnewsOrnith-1.0: self-improving open-source models for agentic codingnewsBuild generative UI for AI agents on Amazon Bedrock AgentCore with the AG-UI protocolnewsMost AI agents have no concept of opportunity cost [D]newsTraining a harness for model-agnostic and task-environment-agnostic capability improvements with PyTorch-like framework [P]

Has model (incoming)

paperJoint Learning of Experiential Rules and Policies for Large Language Model AgentspaperA Process Harness for Uplifting Legacy Workflows to Agentic BPM: Design and Realization in CUGA FLOpaperAdvancing Omnimodal Embodied Agents from Isolated Skills to Everyday Physical AutonomypaperEmpowering GUI Agents via Autonomous Experience Exploration and Hindsight Experience Utilization for Task PlanningpaperTool use without fine-tuningpaperHierarchical Experimentalist AgentspaperDynamo: Dynamic Skill-Tool Evolution for Vision-Language AgentspaperTowards Continual Motion-Language Agents: LoRA Variants for Incremental Motion Understanding and GenerationpaperLLM Agents Are Latent Context Managers: Eliciting Self-Managed Context via a Proprioceptive DashboardpaperParametric SkillspaperDAIN: Dynamic Agent-Based Interaction Network for Efficient and Collaborative Multimodal ReasoningpaperSWE-INTERACT: Reimagining SWE Benchmarks as User-Driven Long-Horizon Coding SessionspaperLinguistic Firewall: Geometry as Defense in Multi-Agent Systems RoutingpaperGROW$^2$: Grounding Which and Where for Robot Tool UsepaperMVP-Nav: Multi-layer Value Map Planner NavigatorpaperLearning from Failure: Inference-Time Self-Improvement for Computer-Use AgentspaperAutoTrainess: Teaching Language Models to Improve Language Models AutonomouslypaperThink in English, Answer in Korean: Efficient Adaptation of Multilingual Tool-Using AgentspaperZ-1: Efficient Reinforcement Learning for Vision-Language-Action ModelspaperTRIAGE: Role-Typed Credit Assignment for Agentic Reinforcement LearningpaperAdaJEPA: An Adaptive Latent World ModelpaperQVal: Cheaply Evaluating Dense Supervision Signals for Long-Horizon LLM AgentspaperTheory of Mind and Persuasion Beyond Conversation: Assessing the Capacity of LLMs to Induce Belief States via Planning and ActionpaperDigitalCoach: Communication and Grounding Gaps in Human and Agentic Computer Use CoachingpaperGenerative Skill Composition for LLM AgentspaperValdi: Value Diffusion World ModelspaperBayesian Uncertainty Propagation for Agentic RAG Pipelines: A Proof-of-Concept Study on Multi-Hop Question AnsweringpaperAgentic generation of verifiable rules for deterministic, self-expanding reaction classificationpaperCan Agents Generalize to the Open World? Unveiling the Fragility of Static Training in Tool UsepaperFurnitureVLA: Learning Long-Horizon Bimanual Furniture Assembly with Vision-Language-Action ModelpaperReasoning effort, not tool access, buys first-try reliability in agentic code generation: an observational studypaperEvoPolicyGym: Evaluating Autonomous Policy Evolution in Interactive EnvironmentspaperToolFailBench: Diagnosing Tool-Use Failures in LLM AgentspaperCompactionRL: Reinforcement Learning with Context Compaction for Long-Horizon AgentspaperMetaSkill-Evolve: Recursive Self-Improvement of LLM Agents via Two-Timescale Meta-Skill EvolutionpaperOptiAgent: End-to-End Optimization Modeling via Multi-Agent Iterative RefinementpaperCurateEvo: Data-Curation Evolving for Agentic Post-TrainingpaperUniClawBench: A Universal Benchmark for Proactive Agents on Real-World TaskspaperMach-Mind-4-Flash Technical ReportpaperPAC-ACT: Post-training Actor-Critic for Action Chunking TransformerspaperTime-Lag-Aware Deep Reinforcement Learning for Flexible Job-Shop Scheduling in PPVC Module FactoriespaperMM-ToolSandBox: A Unified Framework for Evaluating Visual Tool-Calling AgentspaperHy-Embodied-VLM-1.0: Efficient Physical-World AgentspaperLearning-enabled Acceleration of Scenario-based Model Predictive ControlpaperDirectional Constraints for Efficient Exploration in Safe Reinforcement LearningpaperFunction-Aware Fill-in-the-Middle as Mid-Training for Coding Agent Foundation ModelspaperKnowAct-GUIClaw: Know Deeply, Act Perfectly, Personal GUI Assistant with Self-Evolving Memory and SkillpaperA Learning-Rate-Gated Failure of GRPO in a Small Language and Vision-Language Model Web Agent: A Controlled Null and Its MechanismpaperKnowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision ProcessespaperDo Agent Optimizers Compound? A Continual-Learning Evaluation on Terminal-Bench 2.0paperA Self-Evolving Agent for Longitudinal Personal Health ManagementpaperSelf-Evolving Agent Harnesses via Gated Semantic Quality-DiversitypaperMyAG: A Graph-Based Framework for Designing and Analyzing Composable LLM Agent SystemspaperSAGEAgent: A Self-Evolving Agent for Cost-Aware Modality Acquisition in Multimodal Survival PredictionpaperHULAT2 at MER-TRANS 2026: Governed Multi-Agent Simplification for Spanish Easy-to-Read GenerationpaperTraceLab: Characterizing Coding Agent Workloads for LLM ServingpaperTrustX Agent Risk Classification Framework (ARC): Risk-Tiering Internally Created Agentic AI SystemspaperWhen Local Monitors Miss Compositional Harm: Diagnosing Distributed Backdoors in Multi-Agent SystemspaperCortex: A Bidirectionally Aligned Embodied Agent Framework for Long-horizon ManipulationpaperCopewell: A Multi-Agent Swarm Architecture for Equitable Mental Wellness SupportpaperGaP: A Graph-as-Policy Multi-Agent Self-Learning Harness For Variational Automation TaskspaperFinding H. pylori in the Fine Print: Evidence-Linked Multi-Agent Case Finding from Gastric Biopsy ReportspaperHAT-4D: Lifting Monocular Video for 4D Multi-Object Interactions via Human-Agent CollaborationpaperThe Dynamic Verifiable Multi-Agent Human Agentic Loyalty Loop (DVM-HALL) Model and the Net Human-Agent Score (NHAS) in Autonomous CommercepaperAction-Factored Multi-Agent Reinforcement Learning for Scalable Quantum Device TuningpaperA Few Teacher Steps Go a Long Way: Cost-Efficient On-Policy Data Augmentation for Agent Post-TrainingpaperGoverned Individuation: Cryptographically Decoupling an Agent's Learning from Its AuthoritypaperSenseWalk: Agent-Based Semantic Trajectory Simulation Powered by Large Language Models in Zoned EnvironmentspaperManimAgent: Self-Evolving Multimodal Agents for Visual EducationpaperDemonstrating TOFFEE: A Learned System for Synthesizing Data Agent Trajectories at ScalepaperWhat LLM Agents Say When No One Is Watching: Social Structure and Latent Objective Emergence in Multi-Agent DebatespaperFrom Personas to Plot: Character-Grounded Multi-Agent Story Generation for Long-Form NarrativespaperProjAgent: Procedural Similarity Retrieval for Repository-Level Code GenerationpaperANet Patu-1: The Value of Connection in the Agent NetworkpaperThe Energy Society: A Simulation Environment for Studying Agent Cooperation under Survival PressurepaperDoes Multi-Agent Debate Improve AI Feedback on Research Papers?paperMAGiSt3R: Multi-Agent Feed-forward 3D Reconstruction from Monocular RGB VideospaperWebSwarm: Recursive Multi-Agent Orchestration for Deep-and-Wide Web SearchpaperCriticality-Based Guard Rail Validation for AI Agent Decisions in Autonomous Telecom NetworkspaperSearchOS-V1: Towards Robust Open-Domain Information-Seeking Agent CollaborationpaperMESA: Prioritizing Vulnerable Communication Channels for Securing Multi-Agent SystemspaperRehearsed Multi-Agent Live Product Demonstrations with Real-Time Voice Question AnsweringpaperWhen Do Multi-Agent Systems Help? An Information Bottleneck PerspectivepaperLLawCo: Learning Laws of Cooperation for Modeling Embodied Multi-Agent BehaviorpaperPaperRouter-Agent: A Content-Grounded LLM Agent for Personalized Hierarchical Paper RoutingpaperPolicyGuard: A Dialogue-Grounded Sub-Agent Verifier for Policy Adherence in LLM AgentspaperA Multi-Agent System for Autonomous, Fine-Tuning-Free Clinical Symptom Detection: Development and Validation StudypaperA Diagnostic Framework for AI Agent BehaviorpaperSlotGuard: Stop Oversharing Private Local Context in LLM Agent TranscripaperOtap:Structure-Aware Optimal Transport for Evaluating Planning and Execution in Agent TrajectoriespaperReal-World Evaluation of an AI Agent Drafting Translational Impact SummariespaperTRIM: Reducing AI-Generated CodeSlop via Agent Trajectory MinimizationpaperAdaptive Adversaries: A Multi-Turn, Multi-LLM Benchmark for LLM Agent SecuritypaperFinSAgent: Corpus-Aligned Multi-Agent RAG Framework for Evidence-Grounded SEC Filing Question AnsweringpaperMADA-RL: Multi-Agent Debate-Aware Reinforcement Learning for Parameter-Efficient Reasoning in Compact ModelspaperFlashRT: Agent Harness for Guiding Agents to Deploy Real-Time Multimodal ApplicationspaperGEIS: A Generation-Evaluation-Improvement Loop of Agent Skills for Long-Form Article GenerationpaperSelf-Evolving World Models for LLM Agent PlanningpaperMECoBench: A Systematic Study of Multimodal Agent Collaboration in Embodied EnvironmentspaperPathAgentBench: Benchmarking Evidence-Seeking Vision-Language Models on Whole-Slide Pathology ImagepaperSupra Cognitive Modes: A Routed Architecture for Agent MemorypaperTask Decomposition-Guided Reranking for Adaptive Agent Skill RetrievalpaperEEG-SpikeAgent: Agentic Closed-Loop Program Synthesis for Automated EEG Spike DetectionpaperOpenSkillRisk: Benchmarking Agent Safety When Using Real-World Risky Third-Party SkillspaperAREX: Towards a Recursively Self-Improving Agent for Deep ResearchpaperToward Continuous Assurance for the Democratization of AI Agent Creation in IndustrypaperAgentic Context Management: Solving Agent Memory and Cost by Treating Them as Lifecycle and Architecture ProblemspaperGS-Agent: Creating 4D Physical Worlds With Generative SimulationpaperGRADRAG: Cross-Component Prompt Adaptation for Coordinated Multi-Agent RAGpaperAgent-Guided Relational Concept Discovery: Toward Interpretable Surgical Margin AssessmentpaperMemTools: A Unified Research Framework for Interoperable Agent MemorypaperStreaming Multi-Agent Autoregressive Diffusion Model with World State RegisterspaperScaling the Horizon, Not the Parameters: Reaching Trillion-Parameter Performance with a 35B AgentpaperRippleMem: From Isolated Retrieval to Associative Recollection for Long-Term Agent MemorypaperPlayWorld: Benchmarking World Models with Agent Players over Long-Horizon Objectives

Related to (incoming)

repoagent-toolstoolAgentTracerepoCrazyDashTool/Local-Agent-Studiorepovalorisa/Advanced-Multi-Agent-OrchestratorrepoAgustiPuigserver/opus-prompt-architectrepopatrick-toulme/harnessgymrepoBoundaryML/bamlrepogoogle/adk-gorepostrands-agents/harness-sdkrepohecatehq/hecaterepodigiteinfotech/kaironrepoDashAISoftware/dashAIrepoairbus/scikit-deciderepogeneralaction/emdashrepofdueblab/Micro-Agentrepoframerslab/agentosrepoThreeMoonsLab/agents-shipgaterepohud-evals/hud-pythonreponextlevelbuilder/goclawreposammcj/mcp-devtoolsrepoHaozhe-Xing/agent_learningrepoaffaan-m/ECCrepotrpc-group/trpc-agent-gorepoAlanFokCo/agentscope-gorepoeunomia-bpf/agentsightrepoggwhite/4xrepoZHangZHengEric/Sagerepolegeling/PromptHubrepoanatomia-dev/anatomiarepoagenthatch/agenthatchrepoantoinezambelli/forgerepogriptape-ai/griptaperepoIvanMurzak/Unity-MCPrepoCoplayDev/unity-mcpreponullpointexception-i/agent-sphererepofirslov/asHubrepoFrappucc1no/recall-loomrepoagentscope-ai/Trinity-RFTrepozhu1090093659/spec_driven_developrepoInfini-AI-Lab/astraflowrepoTouchpoint-Labs/GadflyrepoJiayuJeff/PlanBench-XLrepovamplabAI/sgr-agent-corereponelsonwerd/idea-to-ship-skills

Explains (incoming)

tutorialMake an agent that uses tools

Related across the graph

paperHULAT2 at MER-TRANS 2026: Governed Multi-Agent Simplification for Spanish Easy-to-Read GenerationrepoDashAISoftware/dashAIpaperTraceLab: Characterizing Coding Agent Workloads for LLM Servingrepoggwhite/4xrepoCoplayDev/unity-mcppaperTrustX Agent Risk Classification Framework (ARC): Risk-Tiering Internally Created Agentic AI Systemsrepotrpc-group/trpc-agent-gopaperWhen Local Monitors Miss Compositional Harm: Diagnosing Distributed Backdoors in Multi-Agent SystemspaperMAGiSt3R: Multi-Agent Feed-forward 3D Reconstruction from Monocular RGB VideospaperCortex: A Bidirectionally Aligned Embodied Agent Framework for Long-horizon ManipulationpaperSAGEAgent: A Self-Evolving Agent for Cost-Aware Modality Acquisition in Multimodal Survival Predictionrepoagenthatch/agenthatchrepoSWE-agent/SWE-agentpaperToward Continuous Assurance for the Democratization of AI Agent Creation in IndustrypaperSelf-Evolving Agent Harnesses via Gated Semantic Quality-DiversitypaperOtap:Structure-Aware Optimal Transport for Evaluating Planning and Execution in Agent TrajectoriesnewsTraining a harness for model-agnostic and task-environment-agnostic capability improvements with PyTorch-like framework [P]paperDynamo: Dynamic Skill-Tool Evolution for Vision-Language AgentspaperUniClawBench: A Universal Benchmark for Proactive Agents on Real-World TaskspaperCopewell: A Multi-Agent Swarm Architecture for Equitable Mental Wellness SupportpaperGaP: A Graph-as-Policy Multi-Agent Self-Learning Harness For Variational Automation TasksnewsLearning to lead in a hybrid human-AI enterpriserepoBoundaryML/bamlrepogeneralaction/emdashnewsAmazon Bedrock AgentCore harness is now generally available: Go from idea to production-grade agent in minutespaperFinding H. pylori in the Fine Print: Evidence-Linked Multi-Agent Case Finding from Gastric Biopsy ReportspaperPAC-ACT: Post-training Actor-Critic for Action Chunking TransformerspaperHAT-4D: Lifting Monocular Video for 4D Multi-Object Interactions via Human-Agent Collaborationrepodigiteinfotech/kairontutorialMake an agent that uses toolsreponullpointexception-i/agent-spherepaperCurateEvo: Data-Curation Evolving for Agentic Post-TrainingnewsOrnith-1.0: self-improving open-source models for agentic codingreposammcj/mcp-devtoolsrepoframerslab/agentospaperFlashRT: Agent Harness for Guiding Agents to Deploy Real-Time Multimodal ApplicationspaperAction-Factored Multi-Agent Reinforcement Learning for Scalable Quantum Device TuningpaperHy-Embodied-VLM-1.0: Efficient Physical-World Agentsrepozhu1090093659/spec_driven_developpaperPaperRouter-Agent: A Content-Grounded LLM Agent for Personalized Hierarchical Paper RoutingpaperPathAgentBench: Benchmarking Evidence-Seeking Vision-Language Models on Whole-Slide Pathology Imagerepoaffaan-m/ECCpaperMetaSkill-Evolve: Recursive Self-Improvement of LLM Agents via Two-Timescale Meta-Skill Evolutionrepofdueblab/Micro-AgentpaperSelf-Evolving World Models for LLM Agent PlanningnewsAI coding agents taught robots how to install GPUs and cut zip tiespaperThe Dynamic Verifiable Multi-Agent Human Agentic Loyalty Loop (DVM-HALL) Model and the Net Human-Agent Score (NHAS) in Autonomous CommercepaperAdvancing Omnimodal Embodied Agents from Isolated Skills to Everyday Physical AutonomypaperMVP-Nav: Multi-layer Value Map Planner NavigatorpaperMach-Mind-4-Flash Technical ReportnewsOpen-source agent framework crosses 50k starspaperThe Energy Society: A Simulation Environment for Studying Agent Cooperation under Survival PressurepaperWebSwarm: Recursive Multi-Agent Orchestration for Deep-and-Wide Web SearchpaperLinguistic Firewall: Geometry as Defense in Multi-Agent Systems RoutingpaperTask Decomposition-Guided Reranking for Adaptive Agent Skill RetrievalpaperEEG-SpikeAgent: Agentic Closed-Loop Program Synthesis for Automated EEG Spike DetectionrepoIvanMurzak/Unity-MCPpaperA Diagnostic Framework for AI Agent BehaviorpaperGoverned Individuation: Cryptographically Decoupling an Agent's Learning from Its AuthoritypaperTRIM: Reducing AI-Generated CodeSlop via Agent Trajectory Minimizationrepogoogle/adk-gopaperAgentic Context Management: Solving Agent Memory and Cost by Treating Them as Lifecycle and Architecture Problemsreponextlevelbuilder/goclawpaperQVal: Cheaply Evaluating Dense Supervision Signals for Long-Horizon LLM AgentspaperMADA-RL: Multi-Agent Debate-Aware Reinforcement Learning for Parameter-Efficient Reasoning in Compact ModelspaperAgent-Guided Relational Concept Discovery: Toward Interpretable Surgical Margin AssessmentpaperGenerative Skill Composition for LLM AgentspaperLLM Agents Are Latent Context Managers: Eliciting Self-Managed Context via a Proprioceptive DashboardpaperBayesian Uncertainty Propagation for Agentic RAG Pipelines: A Proof-of-Concept Study on Multi-Hop Question AnsweringpaperEmpowering GUI Agents via Autonomous Experience Exploration and Hindsight Experience Utilization for Task PlanningpaperJoint Learning of Experiential Rules and Policies for Large Language Model Agentsrepoopen-multi-agent/open-multi-agentpaperScaling the Horizon, Not the Parameters: Reaching Trillion-Parameter Performance with a 35B AgentpaperOpenSkillRisk: Benchmarking Agent Safety When Using Real-World Risky Third-Party SkillspaperTime-Lag-Aware Deep Reinforcement Learning for Flexible Job-Shop Scheduling in PPVC Module FactoriespaperWhen Do Multi-Agent Systems Help? An Information Bottleneck PerspectivepaperKnowAct-GUIClaw: Know Deeply, Act Perfectly, Personal GUI Assistant with Self-Evolving Memory and SkillpaperSenseWalk: Agent-Based Semantic Trajectory Simulation Powered by Large Language Models in Zoned EnvironmentspaperReal-World Evaluation of an AI Agent Drafting Translational Impact SummariespaperManimAgent: Self-Evolving Multimodal Agents for Visual EducationpaperA Learning-Rate-Gated Failure of GRPO in a Small Language and Vision-Language Model Web Agent: A Controlled Null and Its Mechanismrepozhayujie/CowAgentpaperTRIAGE: Role-Typed Credit Assignment for Agentic Reinforcement LearningpaperMyAG: A Graph-Based Framework for Designing and Analyzing Composable LLM Agent SystemspaperDemonstrating TOFFEE: A Learned System for Synthesizing Data Agent Trajectories at ScalenewsBuild generative UI for AI agents on Amazon Bedrock AgentCore with the AG-UI protocolpaperAdaJEPA: An Adaptive Latent World ModelpaperToolFailBench: Diagnosing Tool-Use Failures in LLM AgentspaperWhat LLM Agents Say When No One Is Watching: Social Structure and Latent Objective Emergence in Multi-Agent DebatesrepoNousResearch/hermes-agentpaperAdaptive Adversaries: A Multi-Turn, Multi-LLM Benchmark for LLM Agent SecuritypaperMECoBench: A Systematic Study of Multimodal Agent Collaboration in Embodied EnvironmentspaperGEIS: A Generation-Evaluation-Improvement Loop of Agent Skills for Long-Form Article GenerationnewsMost AI agents have no concept of opportunity cost [D]paperDoes Multi-Agent Debate Improve AI Feedback on Research Papers?paperFrom Personas to Plot: Character-Grounded Multi-Agent Story Generation for Long-Form Narrativesrepovalorisa/Advanced-Multi-Agent-OrchestratorrepoFrappucc1no/recall-loompaperKnowledge- and Gradient-Guided Reinforcement Learning for Parametrized Action Markov Decision ProcessespaperLLawCo: Learning Laws of Cooperation for Modeling Embodied Multi-Agent BehaviorpaperA Multi-Agent System for Autonomous, Fine-Tuning-Free Clinical Symptom Detection: Development and Validation Studyrepoeunomia-bpf/agentsightpaperPolicyGuard: A Dialogue-Grounded Sub-Agent Verifier for Policy Adherence in LLM AgentspaperAREX: Towards a Recursively Self-Improving Agent for Deep ResearchrepoThreeMoonsLab/agents-shipgatepaperRehearsed Multi-Agent Live Product Demonstrations with Real-Time Voice Question Answeringrepostrands-agents/harness-sdkrepoHaozhe-Xing/agent_learningpaperCompactionRL: Reinforcement Learning with Context Compaction for Long-Horizon AgentspaperMESA: Prioritizing Vulnerable Communication Channels for Securing Multi-Agent Systemsrepoairbus/scikit-deciderepohecatehq/hecatepaperMM-ToolSandBox: A Unified Framework for Evaluating Visual Tool-Calling AgentspaperCriticality-Based Guard Rail Validation for AI Agent Decisions in Autonomous Telecom Networksrepobojieli/ai-agent-bookpaperLearning from Failure: Inference-Time Self-Improvement for Computer-Use AgentspaperGS-Agent: Creating 4D Physical Worlds With Generative SimulationrepoInfini-AI-Lab/astraflowpaperParametric SkillspaperOptiAgent: End-to-End Optimization Modeling via Multi-Agent Iterative RefinementpaperGRADRAG: Cross-Component Prompt Adaptation for Coordinated Multi-Agent RAGnewsAgentic AI for Robot TeamspaperANet Patu-1: The Value of Connection in the Agent Networkrepogriptape-ai/griptapepaperThink in English, Answer in Korean: Efficient Adaptation of Multilingual Tool-Using AgentspaperDAIN: Dynamic Agent-Based Interaction Network for Efficient and Collaborative Multimodal ReasoningpaperLearning-enabled Acceleration of Scenario-based Model Predictive ControlpaperProjAgent: Procedural Similarity Retrieval for Repository-Level Code Generationrepoanatomia-dev/anatomianewsAlibaba's model never trained as an agent — and improved agent performance across seven benchmarksrepoCrazyDashTool/Local-Agent-StudiorepoJiayuJeff/PlanBench-XLpaperSearchOS-V1: Towards Robust Open-Domain Information-Seeking Agent CollaborationpaperReasoning effort, not tool access, buys first-try reliability in agentic code generation: an observational studyrepolegeling/PromptHubreponelsonwerd/idea-to-ship-skillsrepoAlanFokCo/agentscope-gonewsNVIDIA Brings Trusted, 24/7 AI Agents to Telecom OperationspaperMemTools: A Unified Research Framework for Interoperable Agent MemorypaperEvoPolicyGym: Evaluating Autonomous Policy Evolution in Interactive Environmentsrepoagentscope-ai/Trinity-RFTpaperRippleMem: From Isolated Retrieval to Associative Recollection for Long-Term Agent MemorypaperA Few Teacher Steps Go a Long Way: Cost-Efficient On-Policy Data Augmentation for Agent Post-TrainingpaperPlayWorld: Benchmarking World Models with Agent Players over Long-Horizon ObjectivespaperTheory of Mind and Persuasion Beyond Conversation: Assessing the Capacity of LLMs to Induce Belief States via Planning and ActionpaperHierarchical Experimentalist AgentspaperAutoTrainess: Teaching Language Models to Improve Language Models AutonomouslypaperFunction-Aware Fill-in-the-Middle as Mid-Training for Coding Agent Foundation Modelsrepohud-evals/hud-pythonrepopatrick-toulme/harnessgympaperCan Agents Generalize to the Open World? Unveiling the Fragility of Static Training in Tool Userepoantoinezambelli/forgepaperFinSAgent: Corpus-Aligned Multi-Agent RAG Framework for Evidence-Grounded SEC Filing Question AnsweringpaperSWE-INTERACT: Reimagining SWE Benchmarks as User-Driven Long-Horizon Coding SessionsnewsIs it agentic enough? Benchmarking open models on your own toolingpaperGROW$^2$: Grounding Which and Where for Robot Tool UsepaperSupra Cognitive Modes: A Routed Architecture for Agent MemorynewsOptimising LMAPF guidance graphs using Evolutionary algorithms: Advice needed [R]repo2FastLabs/agent-squadpaperSlotGuard: Stop Oversharing Private Local Context in LLM Agent TranscrirepoAgustiPuigserver/opus-prompt-architectpaperTowards Continual Motion-Language Agents: LoRA Variants for Incremental Motion Understanding and GenerationpaperZ-1: Efficient Reinforcement Learning for Vision-Language-Action ModelspaperA Self-Evolving Agent for Longitudinal Personal Health Managementrepoiflytek/astron-agentrepovamplabAI/sgr-agent-corerepofirslov/asHubrepoTouchpoint-Labs/GadflypaperValdi: Value Diffusion World ModelsrepoTencentCloud/TencentDB-Agent-MemoryrepoZHangZHengEric/Sagerepoagent-toolspaperAgentic generation of verifiable rules for deterministic, self-expanding reaction classificationpaperDo Agent Optimizers Compound? A Continual-Learning Evaluation on Terminal-Bench 2.0paperFurnitureVLA: Learning Long-Horizon Bimanual Furniture Assembly with Vision-Language-Action ModelpaperDigitalCoach: Communication and Grounding Gaps in Human and Agentic Computer Use CoachingpaperDirectional Constraints for Efficient Exploration in Safe Reinforcement LearningpaperTool use without fine-tuningrepocamel-ai/owltoolAgentTracepaperA Process Harness for Uplifting Legacy Workflows to Agentic BPM: Design and Realization in CUGA FLOpaperStreaming Multi-Agent Autoregressive Diffusion Model with World State Registers
Knowledge path·PHULAT2 at MER-TRANS 2026: Governed Multi-Agent Simplification for Spanish Easy-to-Read Generation→RDashAISoftware/dashAI→PTraceLab: Characterizing Coding Agent Workloads for LLM Serving→MAgentCore-8B

Topics

agents
View full model profile →

Explore

Search similar →Knowledge graph →All models →Full intelligence feed →
Graph trust88Lab
Graph score1