Tag
This paper critiques the reliance on Perspective API for toxicity measurement and argues for community-owned measurement infrastructure in NLP and LLM evaluation, releasing scores for 5.9 million text snippets to facilitate research after the API's shutdown.