Tag
This paper presents an LLM pipeline that converts clinical interview audio into transcripts, maps them to the ten MADRS depression items, estimates severity, and flags problematic ratings. Evaluation on real clinical interviews shows a strong correlation of 0.867 with expert ratings, offering interpretable support for depression assessment in clinical trials.