newsReddit r/MachineLearningTrust 72 · CommunityPublished 1mo agoLive · 1mo ago
Evaluating long-term memory limits in stateless LLM chatbots — feedback needed [D]
Hi all, I’m working on a research project exploring how stateless LLM-based chatbots handle long conversations and whether important earlier information is still reliably retained over time. My idea is to: Run a chatbot using an LLM API without any external memory system Introduce key facts early in a long conversation Continue with many unrelated messages (hundreds of turns) Later test whether the model can
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.
- LinkedLinked via unknownlas7/memharness →
- LinkedLinked via unknownNoshkoto/Noshy →
- LinkedLinked via unknownSelective Memory Retention for Long-Horizon LLM Agents →
- LinkedLinked via unknownManufactured Confidence: How Memory Consolidation Turns Hearsay into Confident Facts →
- LinkedLinked via unknownForensic Trajectory Signatures for Agent Memory Poisoning Detection →
- LinkedLinked via unknownECHO: Prune to act, trace to learn with selective turn memory in agentic RL →
Covers
Covers (incoming)
paperSelective Memory Retention for Long-Horizon LLM AgentspaperManufactured Confidence: How Memory Consolidation Turns Hearsay into Confident FactspaperForensic Trajectory Signatures for Agent Memory Poisoning DetectionpaperWhen the Database Fails: Prompting LLM Dialogue Agents for Safe Recovery in Task-Oriented DialoguepaperECHO: Prune to act, trace to learn with selective turn memory in agentic RLpaperTheory of Mind and Persuasion Beyond Conversation: Assessing the Capacity of LLMs to Induce Belief States via Planning and ActionpaperMessage Passing Enables Efficient ReasoningpaperMemSyco-Bench: Benchmarking Sycophancy in Agent MemorypaperTowards Developing a Multimodal Chat Assistant for University Stakeholders: RAG-based Approachrepoplur-ai/plurrepoTeleAI-UAGI/Awesome-Agent-Memoryrepobasicmachines-co/basic-memoryrepoChatLunaLab/chatlunarepoplastic-labs/honchorepomjason/longpaperAgenticSTS: A Bounded-Memory Testbed for Long-Horizon LLM AgentsrepoMininglamp-OSS/octo-smart-summaryrepoicarito/gtk-llm-chatrepopsinetron/echoes-vault-opencoderepoNirDiamant/Agent_Memory_TechniquespaperDoomed from the Start: Early Abort of LLM Agent Episodes via a Recall-Controlled Probe CascadepaperImproving Ad-hoc Search Effectiveness for Conversational Information Retrieval via Model Mergingrepoabhisadineni/ChatBotrepoericmjl/llamabotreporasbt/LLMs-from-scratchrepoDaoyuanLi2816/pairjudgepaperEvaluating Large Language Models on Misconceptions in Multi-Turn Medical ConversationspaperMemOps: Benchmarking Lifecycle Memory Operations in Long-Horizon Conversationsrepofundamental-research-labs/langdagrepoaiming-lab/SimpleMemreposchaeferms/chat-historyreposchaefer-services/chat-historypaperOne More Turn, Less Regret: A Regret-Based Multi-Turn Benchmark for LLMs' Clarification PoliciespaperRUMBA: Russian User Memory Benchmarkrepochenxiachan/thoughtdagrepommr710/nightmux
Related across the graph
repobasicmachines-co/basic-memoryrepoNirDiamant/Agent_Memory_TechniquespaperEvaluating Large Language Models on Misconceptions in Multi-Turn Medical ConversationspaperAgenticSTS: A Bounded-Memory Testbed for Long-Horizon LLM AgentsrepoMininglamp-OSS/octo-smart-summaryreporasbt/LLMs-from-scratchrepoericmjl/llamabotrepomjason/longrepoplastic-labs/honchopaperMemOps: Benchmarking Lifecycle Memory Operations in Long-Horizon ConversationsrepoDaoyuanLi2816/pairjudgepaperMessage Passing Enables Efficient Reasoningrepoicarito/gtk-llm-chatpaperRUMBA: Russian User Memory BenchmarkpaperOne More Turn, Less Regret: A Regret-Based Multi-Turn Benchmark for LLMs' Clarification PoliciespaperManufactured Confidence: How Memory Consolidation Turns Hearsay into Confident FactspaperECHO: Prune to act, trace to learn with selective turn memory in agentic RLrepopsinetron/echoes-vault-opencoderepoaiming-lab/SimpleMemrepoChatLunaLab/chatlunarepochenxiachan/thoughtdagpaperTowards Developing a Multimodal Chat Assistant for University Stakeholders: RAG-based ApproachpaperImproving Ad-hoc Search Effectiveness for Conversational Information Retrieval via Model Mergingrepofundamental-research-labs/langdagpaperDoomed from the Start: Early Abort of LLM Agent Episodes via a Recall-Controlled Probe Cascaderepommr710/nightmuxpaperTheory of Mind and Persuasion Beyond Conversation: Assessing the Capacity of LLMs to Induce Belief States via Planning and ActionpaperSelective Memory Retention for Long-Horizon LLM Agentsrepolas7/memharnessrepoabhisadineni/ChatBotpaperMemSyco-Bench: Benchmarking Sycophancy in Agent MemorypaperForensic Trajectory Signatures for Agent Memory Poisoning Detectionreposchaeferms/chat-historyrepoTeleAI-UAGI/Awesome-Agent-Memoryrepoplur-ai/plurrepoNoshkoto/NoshypaperWhen the Database Fails: Prompting LLM Dialogue Agents for Safe Recovery in Task-Oriented Dialoguereposchaefer-services/chat-history
