NeurIPS used uncalibrated AI detector for desk rejections [D]
Summary
A submission was desk-rejected from NeurIPS based on an uncalibrated AI detector (Pangram), raising concerns about circularity in the review process and unvalidated false-positive rates on the target distribution.
Similar Articles
Top AI conference uses AI detector to reject papers for allegedly being written by AI
NeurIPS 2026 used a proprietary AI-text detector to desk-reject papers for alleged AI policy violations without validating it on the target distribution; the same detector later flagged conference chairs' own papers as likely AI-written.
NeurIPS 2026 AI-generated reviews [D]
Discussion about the use of AI-generated reviews at NeurIPS 2026, including concerns over prompt injection and lack of consequences for reviewers using LLMs without proper oversight.
@rohanpaul_ai: The AI detector industry should be very uncomfortable reading this MIT report. The strongest institutional rejections y…
An MIT report strongly advises against relying on AI detectors, citing risks such as arms races with AI humanizers, false positives harming students, and unfair impacts on non-native English speakers and neurodivergent individuals.
Why AI Detection Fails for Academic Integrity
This paper evaluates commercial AI detectors in academic settings, finding high false-positive rates on AI-assisted human writing and near-total evasion via humanizers, concluding detector scores should not be standalone misconduct evidence.
NeurIPS 2026 Reviewer: AI-Generated Rebuttals (and Paper) [D]
A NeurIPS reviewer reports encountering a paper and rebuttals that appear entirely LLM-generated, expressing frustration and seeking advice on how to evaluate such submissions.