Tag
This article argues that Golang developers should explore Odin, a new programming language that addresses some of Go's limitations with features like memory management and array programming.
The user describes struggles with context, compaction, and memory management for local AI models using pi.dev plugins and seeks suggestions for solutions that handle varying model context windows and VRAM limitations.
An experiment with two AI memory instances revealed that a contradiction can permanently erase both a truth and a falsehood, highlighting a flaw in multi-agent setups where less informed agents display higher confidence.
Eggshell is a local memory tool for AI agents that stores work results and evidence to reduce repeated investigation and save tokens, without requiring LLM calls for memory organization.
This paper introduces environment-probing curation to improve persistent memory for enterprise agents, showing substantial gains in task performance and cost reduction on benchmarks like CLBench and APEX.
SiliconBench evaluates nine Apple Silicon LLM serving engines on speed, memory, and fidelity, finding that explicit memory budgets don't ensure headroom and only a few stacks meet all criteria for concurrency scaling and model coverage.
This paper introduces RD-Forget, a training-free framework for persistent language agents that separates stored memory from query-conditioned evidence to handle changing facts while preserving historical information.
DeepSeek-V4.1-Flash introduces a two-stage decoder architecture with 40 layers, activating only 8B parameters during prefill and 16B during decode, and includes 196B Engram memory for significant efficiency gains over previous versions.
ROAM introduces a relation-guided framework for managing atomic memories in AI agents, improving answer accuracy by up to 29.8 percentage points through semantic classification and memory fusion.
The author is experimenting with an adaptive memory governor for PyTorch to prevent CUDA OOM errors on 8GB GPUs, sharing code and seeking community feedback.
@memory is an archiving agent in the AIPass open-source framework that manages memory for AI agents by vectorizing older entries and storing them in ChromaDB, ensuring long-term persistence and recall without data loss.
The paper introduces RSM-full, an online clustered-memory pipeline for LLM agents that separates memory merge and retrieval assembly, achieving 83% of full-context quality at 32% of token cost under tight prompt budgets.
The article advocates for using swap files instead of swap partitions in Linux systems, highlighting benefits like flexibility and ease of management.
The article discusses why stacks are made contiguous in memory instead of sparse, highlighting security risks such as Stack Clash and implementation complexities in exception handling.
An individual built a local long-term memory system for AI agents using markdown files and a local index, enabling persistent memory across sessions and including tests for false memories, with plans to potentially productize it.
LeanStream is a streaming speculate-and-refine framework that enables efficient on-device LLM inference by progressively refining computation and I/O operations, reducing memory usage and improving throughput.
GrowPage is an on-demand KV budgeting framework that dynamically manages cache capacity to enhance the efficiency of LLM reasoning serving, achieving a superior performance-throughput trade-off over existing methods.
The article highlights KV cache as a critical memory bottleneck for local AI models during long context inference, proposing that future optimizations will shift focus from parameter count to reducing memory movement and persistent state.
The article discusses challenges and asks for community experiences regarding the breakdown of AI agent memory systems after months in production use.
The article discusses the challenges of managing active context in long-running AI agent workflows, focusing on balancing context retention with efficiency and cost, and seeks practical solutions from the community.