BEAM benchmarks

Reddit r/AI_Agents News

Summary

Midas achieves 0.56 recall@k on BEAM 100K and 0.51 on BEAM 500K with zero LLM calls and zero cost, demonstrating efficient long-term memory for agents.

Today we ran our first benchmark with Midas on BEAM, one of the most important long-term memory benchmarks for agents. Midas reached 0.56 recall@k on BEAM 100K and 0.51 on BEAM 500K, with 0 LLM calls, $0 API spend, and 0 data egress. 1M and 10M tiers are next. My aim is learn from hindsight and other projects to keep improving Midas while still being local-first 0$ cost. What do you think? Would it be possible to get to that level?
Original Article

Similar Articles