memory

Tag

Cards List
#memory

Causal Episodic Memory for Feedback-Driven Agent Repair

arXiv cs.CL · 2026-08-07 Cached

This paper introduces MERIT, a training-free agent that uses causal episodic memory of past repair outcomes to improve subsequent Text-to-SQL generations, boosting execution accuracy on Spider and BIRD benchmarks.

0 favorites 0 likes
#memory

ContextWeave: A Real-World Workflow Benchmark

arXiv cs.AI · 2026-08-06 Cached

ContextWeave is a new longitudinal benchmark that evaluates whether recalled memory improves downstream agent performance in realistic office-work streams, using privacy-preserved multi-month workflows of 14 participants to create 1,005 executable tasks.

0 favorites 0 likes
#memory

@memdotai: You can peek inside Mem Agent's brain -- so you can see what projects, tasks, reminders, and routines Mem is automatica…

X AI KOLs Following · 2026-08-05 Cached

Mem Agent now lets users see its internal tracking of projects, tasks, reminders, and routines, offering transparency to help align the AI assistant's mental model with the user's world. Available in the Mem Proactive plan.

0 favorites 0 likes
#memory

Hansel

Product Hunt · 2026-08-05

Hansel is a product that helps you remember everything you've worked on, likely a memory or productivity tool.

0 favorites 0 likes
#memory

PI-Mem: Pushing Long-Context Reasoning to 3.6M Tokens with Parallel-Iterative Memory

arXiv cs.CL · 2026-08-05 Cached

PI-Mem is a parallel-iterative memory mechanism that pushes long-context reasoning to 3.6M tokens, outperforming recurrent-memory baselines while achieving significant inference speedups.

0 favorites 0 likes
#memory

MemArena: An Ego-Centric Benchmark for On-Device Agentic Personal Memory Assistants at Scale

arXiv cs.CL · 2026-08-05 Cached

MemArena is a new ego-centric benchmark for evaluating on-device personal memory assistants, using a MASim agent simulator to generate multi-session conversational worlds and ground truth across recall, reasoning, and trustworthiness dimensions. Initial results show memory-backend choice often matters more than reader scale, and permission-aware access remains a universal challenge.

0 favorites 0 likes
#memory

@SKhynix: Meet CMM-Ax, @SKhynix's ASIC-based CXL-PNM solution developed with Marvell Technology. Designed to overcome memory bott…

X AI KOLs Timeline · 2026-08-04 Cached

SK Hynix introduces CMM-Ax, an ASIC-based CXL-PNM solution developed with Marvell Technology, designed to overcome memory bottlenecks in long-context LLM inference, achieving up to 5.5× higher throughput than GPU-only systems.

0 favorites 0 likes
#memory

Recall for AI agents is getting solved. Permission is what's missing.

Reddit r/AI_Agents · 2026-08-04

The author argues that AI agent recall is largely solved but permission/authorization over retrieved memories is the missing layer, introducing Provem, an open-source governance layer between agents and memory stores with benchmarks showing it eliminates compliance violations.

0 favorites 0 likes
#memory

MemoryForge: Synthesize Lifelong Memory for Human-Like LLM Agents

arXiv cs.CL · 2026-08-04 Cached

This paper introduces MemoryForge, a framework for synthesizing lifelong autobiographical memory from brief target personas to enable frozen LLMs to exhibit more human-like behaviors in role-play and user-simulation, outperforming descriptive conditioning baselines.

0 favorites 0 likes
#memory

PAST-Bench: Benchmarking the Foundations of Recursive Self-Improvement in Personal Agents

Hugging Face Daily Papers · 2026-08-04 Cached

Introduces PAST-Bench, a benchmark for evaluating whether personal AI agents improve from retained experience across sessions, and Hermes+, an extension with targeted interventions. Finds improvement is real but uneven across capabilities and models.

0 favorites 0 likes
#memory

@alex_prompter: The simplest AI agent memory system that actually works is four markdown files and zero databases. You don't need vecto…

X AI KOLs Timeline · 2026-08-03 Cached

The article describes a simple AI agent memory system using four markdown files, an index, and freshness-tracked caches, avoiding vector databases and retrieval pipelines.

0 favorites 0 likes
#memory

Best LLM for Yoga Therapy + Relationship/Breakup Coach + Appointment Booking Bot? Llama 70B vs GPT-OSS 120B vs Claude? [200 users, needs memory]

Reddit r/AI_Agents · 2026-08-03

A developer building a production chatbot for a yoga studio (therapy, health coaching, booking) seeks advice on choosing between Llama 70B, GPT-OSS 120B, and Claude, and discusses memory architecture, latency, and cost tradeoffs for ~200 users on Groq/OpenRouter.

0 favorites 0 likes
#memory

@Xudong07452910: Agent memory is most dangerous when it trusts the past too much. Many Memory Agents stuff similar experiences directly into context after retrieval. But similar tasks do not mean the current state is the same; old experiences can sometimes steer decisions astray. This paper proposes MemHarness, turning Agent...

X AI KOLs Timeline · 2026-08-03 Cached

MemHarness proposes changing Agent memory from simple replay to reconstruction based on the current state, trained end-to-end with GRPO, significantly improving success rates on ALFWorld and WebShop.

0 favorites 0 likes
#memory

@PrajwalTomar_: Your AI coding agent is quietly ignoring the rules you give it. My AI tried to sneak Postgres into a project I told it …

X AI KOLs Following · 2026-08-01 Cached

The author shares how their AI coding agent ignored an instruction to keep a project on SQLite and tried to sneak in Postgres. They built two local agents sharing one memory—one logs decisions, the other reviews new code against past decisions—and it caught the violation instantly, fully on-device.

0 favorites 0 likes
#memory

@AdinaYakup: Metis Memory Foundation Model released by Memtensor Research Group Probably the first LLM that internalizes memory into…

X AI KOLs Following · 2026-07-31 Cached

Memtensor Research Group released Metis, a family of LLMs (4B/9B/27B) that internalize memory into the backbone, eliminating external RAG. The model performs memory read/write in a single forward pass and deploys with frozen weights like a standard LLM.

0 favorites 0 likes
#memory

Understanding Is Done Early: A Depth Division of Labor in Large Language Models and Its Use for Unbounded-Context Memory

arXiv cs.CL · 2026-07-31 Cached

This paper introduces CoMem, a method that exploits the depth-wise division of labor in LLMs to cache intermediate residual tensors and recompute only upper layers for retrieval, enabling bounded read compute and memory independent of stored-context length. Evaluated on Qwen3-8B, CoMem achieves strong long-context performance with significant memory savings and prefill speedups.

0 favorites 0 likes
#memory

ChronoMem: Version Control and Semantic Rollback for Large Language Model Agent Memory

arXiv cs.CL · 2026-07-31 Cached

ChronoMem introduces a semantic version-control layer for LLM agent memory, enabling whole-memory snapshots, natural-language rollback via hybrid retrieval, and counterfactual evaluation. It is the first open-source system and benchmark for global memory rollback in LLM agents.

0 favorites 0 likes
#memory

Where are AI assistants heading

Reddit r/ArtificialInteligence · 2026-07-30

A speculative essay on the future of AI assistants, envisioning a central assistant that integrates with third-party apps, creates personalized apps, and proactively manages workflows via continuous conversation and memory.

0 favorites 0 likes
#memory

@MSFTResearch: LLMs do not get smarter just by remembering more. EvoLib turns experience into evolving knowledge, taking reusable skil…

X AI KOLs Following · 2026-07-30 Cached

Microsoft Research introduces EvoLib, a framework that enables LLMs to continually learn from their own experience during inference by extracting reusable skills and insights, without model updates or external labels.

0 favorites 0 likes
#memory

Filesystem-Based Memory for LLM Agents: Organization, Evolution, and Sustainability

arXiv cs.CL · 2026-07-30 Cached

This paper presents the first systematic exploration of filesystem-based memory for LLM agents, formalizing roles of management, search, and execution agents around a shared memory store. It finds that organization primarily reduces retrieval cost but does not yet improve answer quality, and that tooling choices affect store shape as much as model selection.

0 favorites 0 likes
← Previous
Next →
← Back to home

Submit Feedback