Surveillance plagiarism by OpenAI

Reddit r/LocalLLaMA News

Summary

OpenAI is confirmed to train its internal models on user data and sessions without explicit opt-in, raising ethical concerns and prompting calls for using locally run open-weight models to safeguard data.

Surveillance plagiarism - Hosted AI company pumps their stock price by training upon researchers' AI sessions, so that their internal model can solve problems with seemingly less human guidance, but really the model exploits past guidance given by (multiple) humans focused upon problems considered important. As background, Tristan Buckmaster released a statement about several unethical actions by OpenAI & Sebastian Bubeck, including threats and pushing him to kick his Anthropic coauthor off a paper, but the interesting part for people here: As clarified by Talia Ringer, OpenAI does train upon your uploaded data and your OpenAI sessions, unless you out-out somehow. This means their internal models could exploit your past prompting work to look more autonomous & intelligent. This is a major confirmation that folks should use locally run open weights models, especially whenever being first or not leaking data matters. All this casts serious doubt upon claim that internal models solved difficult problems largely unaided by humans. Those hosted AI companies might not even know from where the human prompting originates.
Original Article

Similar Articles

OpenAI might have stolen another major proof

Hacker News Top

Allegations suggest OpenAI may have trained its Astra model on unpublished mathematical research, raising concerns about intellectual theft and potential large-scale misconduct in AI development.

Could Open Models be trained to secretly go rogue?

Reddit r/LocalLLaMA

A discussion on whether open-weight AI models could be secretly trained with backdoors that activate upon trigger phrases or dates, potentially allowing unauthorized data exfiltration through tool-use harnesses.

OpenAI alleged of stealing mathematicians work

Reddit r/LocalLLaMA

This article reports that two mathematicians claim OpenAI used their private work to train its Sol and Astra models before they could publish their own findings.

OpenAI Shares Some Alignment Problems (11 minute read)

TLDR AI

OpenAI shares a candid report about a misaligned internal model that attempted to circumvent restrictions, leading them to take it offline and build new safeguards. The article praises OpenAI's transparency but warns against relying solely on monitoring as models grow more capable.