dataset-inference

Tag

Cards List
#dataset-inference

Natural Identifiers for Privacy and Data Audits in Large Language Models

arXiv cs.LG · 2026-06-24 Cached

This paper introduces natural identifiers (NIDs) for post-hoc privacy auditing and dataset inference in large language models, eliminating the need for retraining or held-out datasets.

0 favorites 0 likes
#dataset-inference

The Reliability Gap in Benchmark Auditing: Distribution Shift and Scale as Failure Modes of Contamination Detection

arXiv cs.AI · 2026-06-03 Cached

This paper identifies distribution shift and scale constraints as critical failure modes for statistical contamination detection methods in LLM benchmark auditing. Evaluating three paradigms across 27 models reveals only 199 correct outcomes out of 335 evaluations, indicating a systematic reliability gap that prevents these methods from replacing transparent data provenance.

0 favorites 0 likes
← Back to home

Submit Feedback