counterfactual-audio

Tag

Cards List
#counterfactual-audio

HEAR Who Said What: Unlocking Speaker-Attributed Reasoning via Counterfactual Voice Grounding

arXiv cs.CL · 2d ago Cached

This paper introduces HEAR, a benchmark for evaluating speaker-attributed reasoning in speech language models, and presents A2R, a 30B model optimized with counterfactual data to improve performance on multi-speaker tasks.

0 favorites 0 likes
← Back to home

Submit Feedback