Tag
VoxMem is a benchmark for evaluating multimodal memory in Large Audio Language Models, focusing on acoustic evidence types and multi-session memory operations. It reveals significant gaps in current models, such as poor retention of speaker identity and paralinguistic cues compared to semantic content.
Zero-Mem introduces a zero-token memory operation system for LLM agents, eliminating LLM calls and token consumption during memory retrieval by preserving original interaction traces and using deterministic entity-context and temporal structures. It achieves competitive performance on long-memory and long-context QA benchmarks while reducing memory-operation time cost by 57.6%.