preference-discovery

Tag

Cards List
#preference-discovery

APeB: Benchmarking Personalization Ability of Large Language Model Agents

arXiv cs.AI · 2026-07-07 Cached

Introduces APeB, a benchmark for evaluating personalization in LLM agents, focusing on inferring user intent and preferences from raw queries and interaction histories. Finds that current models struggle with early-stage queries and that history-aware refinement can help.

0 favorites 0 likes
← Back to home

Submit Feedback