Tag
This paper introduces methods to quantify how effectively a panel of language models approximates human judgments, using spectral diversity and distribution recovery metrics to measure effective representation.