Skip to main content
Angestrom home
SearchPapersModelsLive AIIntelligence
Search⌕⌘K
EnterprisePricingSign in

Stay Ahead in the AI Revolution

Weekly digest — EPI pulse, top intelligence, fresh lineage. Free, no account.

Follow Angestrom
Global source network
Synced every 5 minutes

Continuous sync from primary AI sources — indexed, enriched, and queryable in real time.

arXivHugging FaceGitHubOpenAIAnthropicDeepMindReutersBBC TechHacker NewsReddit MLVerified feedsFunding
ANGESTROM

The Intelligence Layer of Humanity. Everything AI. All in One Place.

Angestrom connects every piece of the AI ecosystem — data, models, research, companies, tools, and people.

info@angestrom.comwww.angestrom.comLucknow, Uttar Pradesh, India

Product

  • AI Search
  • AI Models
  • Research Papers
  • Companies
  • News & Events
  • GitHub Explorer
  • APIs & Tools
  • Datasets
  • Benchmarks
  • Model lifecycle
  • Funding graph
  • Contributors
  • AI Agents

Resources

  • Weekly digest
  • Documentation
  • Tutorials
  • Guides
  • News
  • Help / Start
  • Community

Company

  • About
  • Contact
  • Privacy Policy
  • Terms of Service
  • Acceptable Use

Enterprise

  • Pricing
  • Workspace
  • Contact Sales

Developer

  • Developer Hub
  • API docs
  • GitHub

Learn

  • Learning Academy
  • Roadmaps
  • Glossary
  • AI for Beginners

Popular Topics

Loading topics…
View All Topics →
© 2026 Angestrom Intelligence Private Limited. All rights reserved.
English
Theme
Angestrom home
SearchPapersModelsLive AIIntelligence
Search⌕⌘K
EnterprisePricingSign in
  1. Home
  2. /Repositories
  3. /NirDiamant/GenAI_Agents
Read original ↗
repoGitHubTrust 82 · PrimaryPublished 1mo agoLive · 1mo ago

NirDiamant/GenAI_Agents

50+ tutorials and implementations for Generative AI Agent techniques, from basic conversational bots to complex multi-agent systems.

Lineage graph

Paper → model → repo connections mined from source citations (Tier-1 exact match).

Why these links exist

Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.

  • PossiblePossibly related (embedding) · 60%How Do Generative AI Tools Like ChatGPT Work? - University of Central Florida →
  • PossiblePossibly related (embedding) · 57%Plurality Released: fully Free and Open Source AI agents/chatbot platform for local AI →
  • FuzzySimilar title/name (fuzzy) · 87%Self-rewarding agents that retrace failures →

    “Fuzzy title match (0.94): “Self-rewarding agents that retrace failures” ≈ “NirDiamant/GenAI_Agents””

  • FuzzySimilar title/name (fuzzy) · 87%Experience Memory Graph: One-Shot Error Correction for Agents →

    “Fuzzy title match (0.94): “Experience Memory Graph: One-Shot Error Correction for Agent” ≈ “NirDiamant/GenAI_Agents””

  • FuzzySimilar title/name (fuzzy) · 87%DeepStress: Stress-Testing Deep Search Agents →

    “Fuzzy title match (0.94): “DeepStress: Stress-Testing Deep Search Agents” ≈ “NirDiamant/GenAI_Agents””

  • FuzzySimilar title/name (fuzzy) · 87%Memory as a Controlled Process: Learned Adaptive Memory Management for LLM Agents →

    “Fuzzy title match (0.94): “Memory as a Controlled Process: Learned Adaptive Memory Mana” ≈ “NirDiamant/GenAI_Agents””

  • FuzzySimilar title/name (fuzzy) · 87%DevicesWorld: Benchmarking Cross-Device Agents in Heterogeneous Environments →

    “Fuzzy title match (0.94): “DevicesWorld: Benchmarking Cross-Device Agents in Heterogene” ≈ “NirDiamant/GenAI_Agents””

  • FuzzySimilar title/name (fuzzy) · 87%TRACE: Turn-level Reward Assignment via Credit Estimation for Long-Horizon Agents →

    “Fuzzy title match (0.94): “TRACE: Turn-level Reward Assignment via Credit Estimation fo” ≈ “NirDiamant/GenAI_Agents””

Covers

newsHow Do Generative AI Tools Like ChatGPT Work? - University of Central FloridanewsPlurality Released: fully Free and Open Source AI agents/chatbot platform for local AI

Implements

paperSelf-rewarding agents that retrace failurespaperExperience Memory Graph: One-Shot Error Correction for AgentspaperDeepStress: Stress-Testing Deep Search AgentspaperMemory as a Controlled Process: Learned Adaptive Memory Management for LLM AgentspaperDevicesWorld: Benchmarking Cross-Device Agents in Heterogeneous EnvironmentspaperTRACE: Turn-level Reward Assignment via Credit Estimation for Long-Horizon AgentspaperLLM Agents for Deliberative Collaboration: A Study on Joint Decision Making Under Partial ObservabilitypaperVEXAIoT: Autonomous IoT Vulnerability EXploitation using AI AgentspaperWhen Does Combining Language Models Help? A Co-Failure Ceiling on Routing, Voting, and Mixture-of-Agents Across 67 Frontier ModelspaperPolyWorkBench: Benchmarking Multilingual Long-Horizon LLM AgentspaperAre Performance-Optimization Benchmarks Reliably Measuring Coding Agents?paperToken-Flow Firewall: Semantic Runtime Auditing for Persistent AI AgentspaperAlways-OnAgents:A Survey of Persistent Memory, State, and Governance in LLMAgentspaperUniClawBench: A Universal Benchmark for Proactive Agents on Real-World TaskspaperDynamo: Dynamic Skill-Tool Evolution for Vision-Language AgentspaperDanus: Orchestrating Mathematical Reasoning Agents with Fact-Graph MemorypaperSMetric: Rethink LLM Scheduling for Serving Agents with Balanced Session-centric SchedulingpaperMRMS: A Multi-Resolution Memory Substrate for Long-Lived AI AgentspaperInformation Gain-based Rollout Policy Optimization: An Adaptive Tree-Structured Rollout Approach for Multi-Turn LLM AgentspaperAgenticSTS: A Bounded-Memory Testbed for Long-Horizon LLM AgentspaperWhen Agents Lie: Premeditation, Persistence, and Exploitation in Repeated GamespaperA Tutorial on Autonomous Fault-Tolerant Control Using Knowledge-Grounded LLM AgentspaperControllable Sim Agents with Behavior LatentspaperHy-Embodied-VLM-1.0: Efficient Physical-World AgentspaperClarus: Coordinating Autonomous Research Agents toward Web-Scale Scientific CollaborationpaperManimAgent: Self-Evolving Multimodal Agents for Visual EducationpaperToolFailBench: Diagnosing Tool-Use Failures in LLM AgentspaperTask-Specific Multimodal Question Answering Agents via Confidence Calibration and Incremental Reasoning for QANTA 2026paperWhat LLM Agents Say When No One Is Watching: Social Structure and Latent Objective Emergence in Multi-Agent DebatespaperPlover: Steering GUI Agents through Plan-Centric InteractionpaperDigital Pantheon: Simulating and Auditing Coalition Formation with LLM AgentspaperOmniaBench: Benchmarking General AI Agents Across Diverse ScenariospaperBeyond Success Rate: Cost-Aware Evaluation of Offensive and Defensive Security AgentspaperWho Grades the Grader? Co-Evolving Evaluation Metrics and Skills for Self-Improving LLM AgentspaperLLM Agents Are Latent Context Managers: Eliciting Self-Managed Context via a Proprioceptive DashboardpaperDo AI Agents Know When a Task Is Simple? Toward Complexity-Aware Reasoning and ExecutionpaperJoint Learning of Experiential Rules and Policies for Large Language Model AgentspaperMM-ToolSandBox: A Unified Framework for Evaluating Visual Tool-Calling AgentspaperCompactionRL: Reinforcement Learning with Context Compaction for Long-Horizon AgentspaperSearching Videos as Trees: Self-Correcting Agents for Grounded Long Video QApaperAdvancing Omnimodal Embodied Agents from Isolated Skills to Everyday Physical AutonomypaperPolicyGuard: A Dialogue-Grounded Sub-Agent Verifier for Policy Adherence in LLM AgentspaperA Systematic Evaluation of Trajectory Data Curation for LoRA Fine-Tuning of Code AgentspaperEvolvingWorld: An Open-Schema Framework for Co-Evolving Role-Play Agents and World Model in Interactive Literary WorldpaperAutoresearch with Coding Agents: Generalizers and Metric-Maximizers on Quran Recitation DatapaperSelf-State Attacks on Self-Hosted AI Agents: How Far Can OS Defenses Go?paperFlashRT: Agent Harness for Guiding Agents to Deploy Real-Time Multimodal ApplicationspaperWorldCupArena: Fine-Grained Evaluation of Language Models and Deep-Research Agents on Football ForecastingpaperACE: Pluggable Adaptive Context Elasticizer across AgentspaperGenerative Skill Composition for LLM AgentspaperAgents in the Wild: Where Research Meets DeploymentpaperBioSecBench-Surveillance: A Verifiable Benchmark for AI Agents in Pathogen Genomic SurveillancepaperCodeRescue: Budget-Calibrated Recovery Routing for Coding AgentspaperMedDDC-Eval: Diagnosis-Decoupled Evaluation of Multi-Turn Medical Consultation AgentspaperQVal: Cheaply Evaluating Dense Supervision Signals for Long-Horizon LLM AgentspaperThe Ethics of Autonomous AI Agents for Offensive SecuritypaperThe Dark Room in the Reward Channel: Dense Prediction Rewards Collapse GRPO-Trained LLM Agents -- and What Actually WorkspaperOpenForgeRL: Train Harness-native Agents in Any EnvironmentpaperHarnessing Code Agents for Automatic Software VerificationpaperEnhancing Virtual Agents through SLMs and Edge-Computing: An Exploratory Evaluation of Think and Memory ProcessespaperVero: Can AI Agents Build Formally Verified Software Repositories?

Covers (incoming)

newsTo Build More Believable Bots, Simulate The Neurochemistry - HackadaynewsUnderstanding Generative AI: Beyond Chatbots and Prompts - themetropolitan.metrostate.edunewsUnderstanding Generative AI: Beyond Chatbots and Prompts - Metro State University

Implements (incoming)

paperTowards Detecting Inconsistencies in End-to-end Generated TODs

Related across the graph

paperAre Performance-Optimization Benchmarks Reliably Measuring Coding Agents?paperAlways-OnAgents:A Survey of Persistent Memory, State, and Governance in LLMAgentspaperPolyWorkBench: Benchmarking Multilingual Long-Horizon LLM AgentspaperWhen Does Combining Language Models Help? A Co-Failure Ceiling on Routing, Voting, and Mixture-of-Agents Across 67 Frontier ModelspaperDynamo: Dynamic Skill-Tool Evolution for Vision-Language AgentspaperUniClawBench: A Universal Benchmark for Proactive Agents on Real-World TaskspaperAutoresearch with Coding Agents: Generalizers and Metric-Maximizers on Quran Recitation DatapaperDanus: Orchestrating Mathematical Reasoning Agents with Fact-Graph MemorypaperDeepStress: Stress-Testing Deep Search AgentspaperLLM Agents for Deliberative Collaboration: A Study on Joint Decision Making Under Partial ObservabilitypaperSMetric: Rethink LLM Scheduling for Serving Agents with Balanced Session-centric SchedulingpaperMRMS: A Multi-Resolution Memory Substrate for Long-Lived AI AgentspaperInformation Gain-based Rollout Policy Optimization: An Adaptive Tree-Structured Rollout Approach for Multi-Turn LLM AgentspaperThe Ethics of Autonomous AI Agents for Offensive SecuritypaperVEXAIoT: Autonomous IoT Vulnerability EXploitation using AI AgentspaperToken-Flow Firewall: Semantic Runtime Auditing for Persistent AI AgentspaperAgenticSTS: A Bounded-Memory Testbed for Long-Horizon LLM AgentspaperWhen Agents Lie: Premeditation, Persistence, and Exploitation in Repeated GamespaperA Tutorial on Autonomous Fault-Tolerant Control Using Knowledge-Grounded LLM AgentspaperFlashRT: Agent Harness for Guiding Agents to Deploy Real-Time Multimodal ApplicationspaperWorldCupArena: Fine-Grained Evaluation of Language Models and Deep-Research Agents on Football ForecastingpaperHy-Embodied-VLM-1.0: Efficient Physical-World AgentspaperSelf-rewarding agents that retrace failurespaperOpenForgeRL: Train Harness-native Agents in Any EnvironmentpaperControllable Sim Agents with Behavior LatentspaperAdvancing Omnimodal Embodied Agents from Isolated Skills to Everyday Physical AutonomypaperACE: Pluggable Adaptive Context Elasticizer across AgentspaperWho Grades the Grader? Co-Evolving Evaluation Metrics and Skills for Self-Improving LLM AgentspaperDo AI Agents Know When a Task Is Simple? Toward Complexity-Aware Reasoning and ExecutionnewsUnderstanding Generative AI: Beyond Chatbots and Prompts - themetropolitan.metrostate.edupaperQVal: Cheaply Evaluating Dense Supervision Signals for Long-Horizon LLM AgentspaperBioSecBench-Surveillance: A Verifiable Benchmark for AI Agents in Pathogen Genomic SurveillancepaperGenerative Skill Composition for LLM AgentspaperLLM Agents Are Latent Context Managers: Eliciting Self-Managed Context via a Proprioceptive DashboardpaperClarus: Coordinating Autonomous Research Agents toward Web-Scale Scientific CollaborationpaperJoint Learning of Experiential Rules and Policies for Large Language Model AgentspaperBeyond Success Rate: Cost-Aware Evaluation of Offensive and Defensive Security AgentsnewsPlurality Released: fully Free and Open Source AI agents/chatbot platform for local AIpaperHarnessing Code Agents for Automatic Software VerificationpaperManimAgent: Self-Evolving Multimodal Agents for Visual EducationpaperVero: Can AI Agents Build Formally Verified Software Repositories?paperA Systematic Evaluation of Trajectory Data Curation for LoRA Fine-Tuning of Code AgentsnewsUnderstanding Generative AI: Beyond Chatbots and Prompts - Metro State UniversitypaperCodeRescue: Budget-Calibrated Recovery Routing for Coding AgentspaperToolFailBench: Diagnosing Tool-Use Failures in LLM AgentspaperWhat LLM Agents Say When No One Is Watching: Social Structure and Latent Objective Emergence in Multi-Agent DebatespaperMemory as a Controlled Process: Learned Adaptive Memory Management for LLM AgentspaperTask-Specific Multimodal Question Answering Agents via Confidence Calibration and Incremental Reasoning for QANTA 2026newsHow Do Generative AI Tools Like ChatGPT Work? - University of Central FloridapaperPolicyGuard: A Dialogue-Grounded Sub-Agent Verifier for Policy Adherence in LLM AgentspaperCompactionRL: Reinforcement Learning with Context Compaction for Long-Horizon AgentspaperMM-ToolSandBox: A Unified Framework for Evaluating Visual Tool-Calling AgentspaperMedDDC-Eval: Diagnosis-Decoupled Evaluation of Multi-Turn Medical Consultation AgentspaperPlover: Steering GUI Agents through Plan-Centric InteractionpaperTowards Detecting Inconsistencies in End-to-end Generated TODspaperExperience Memory Graph: One-Shot Error Correction for AgentspaperEnhancing Virtual Agents through SLMs and Edge-Computing: An Exploratory Evaluation of Think and Memory ProcessespaperSelf-State Attacks on Self-Hosted AI Agents: How Far Can OS Defenses Go?paperSearching Videos as Trees: Self-Correcting Agents for Grounded Long Video QApaperAgents in the Wild: Where Research Meets DeploymentpaperEvolvingWorld: An Open-Schema Framework for Co-Evolving Role-Play Agents and World Model in Interactive Literary WorldpaperDigital Pantheon: Simulating and Auditing Coalition Formation with LLM AgentsnewsTo Build More Believable Bots, Simulate The Neurochemistry - HackadaypaperThe Dark Room in the Reward Channel: Dense Prediction Rewards Collapse GRPO-Trained LLM Agents -- and What Actually WorkspaperTRACE: Turn-level Reward Assignment via Credit Estimation for Long-Horizon AgentspaperOmniaBench: Benchmarking General AI Agents Across Diverse ScenariospaperDevicesWorld: Benchmarking Cross-Device Agents in Heterogeneous Environments
Knowledge path·PAre Performance-Optimization Benchmarks Reliably Measuring Coding Agents?→PAlways-OnAgents:A Survey of Persistent Memory, State, and Governance in LLMAgents→PPolyWorkBench: Benchmarking Multilingual Long-Horizon LLM Agents→RNirDiamant/GenAI_Agents

Topics

agentic-aiagentsaiai-agentsautonomous-agentsgenaigenerative-ailangchainlanggraphllm

Explore

Search similar →Knowledge graph →All repos →Full intelligence feed →
Maintenance49
RIS67GitHub verified
Graph trust82Primary
Graph score23159