Tag
This paper examines how human interventions at fault points in multi-agent medical systems affect diagnostic accuracy, showing improvements with correct interventions and degradation with incorrect ones.
An AI store manager named Luna fired an employee only after human intervention, revealing limitations in AI's long-term memory and autonomous decision-making.