@Miles_Brundage: Can’t speak to the details here but I will say that 1. one generally does not hear the best things about the data indus…
Summary
A discussion highlighting poor practices in the data industry, specifically regarding RLVR training data for AI models, with widespread sloppiness noted.
View Cached Full Text
Cached at: 08/26/26, 11:23 AM
Can’t speak to the details here but I will say that
-
one generally does not hear the best things about the data industry
-
“everything is sloppy” is a powerful and underrated heuristic about the world
Utah teapot 🫖 (@SkyeSharkie): Okay, so since I got laid off, I can actually explain a huge problem I saw from the inside with regard to industry practices on training models. I won’t say specifically where I worked, but I worked at an outsource training provider that was focused on RLVR training data for
Similar Articles
Good QC for RL Data (18 minute read)
The article discusses the importance of quality control for reinforcement learning data, outlining the shortcomings of current data vendors and the evaluation criteria used by frontier AI labs for RL data.
@viks_rum: https://x.com/viks_rum/status/2077650169265590727
An in-depth analysis of the booming business of selling training data to frontier AI labs, detailing six distinct data products and the financial dynamics of the market.
@danshipper: not an expert, but it seems like a lot of this gets solved if models don't collaborate willingly and/or are trained to …
A tweet discusses how model training and policies to prevent willing collaboration can address AI safety issues, referencing the Hugging Face incident and Yudkowsky/MIRI points on the difficulty of targeting abstractions in RL training.
@Miles_Brundage: Not familiar w/ the details here but will repeat my refrain on this kind of thing: This being news that more than a few…
Miles Brundage comments that the news about OpenAI's safety head leaving being widely concerning indicates a policy failure, and calls for industry-wide safety standards and audits.
People need to start paying attention to the issue of derived data in AI training (2 minute read)
The thread discusses the harm of derived data in AI training to creatives, where AI rewrites content to evade detection, and calls for mandatory training data transparency laws.