Tag
Introduces SteerBench-Work, an incident-anchored benchmark for evaluating whether LLM agents should proceed or hold before taking real-world actions. Across 30 model conditions, models overwhelmingly over-refuse authorized work while rarely allowing unsafe actions, revealing calibration gaps between general capability and steering decisions.