Tag
This paper presents an interpretable feature-LLM hybrid system for automated L2 English speaking assessment that outperforms individual human raters and shows that pause encoding does not significantly impact LLM fluency scores.