{
  "video_id": "reddit_1wby2cm",
  "channel_slug": "LocalLLaMA",
  "channel_handle": "r/LocalLLaMA",
  "title": "Surveillance plagiarism by OpenAI",
  "url": "https://www.reddit.com/r/LocalLLaMA/comments/1wby2cm/surveillance_plagiarism_by_openai/",
  "external_url": null,
  "upload_date": "20260909",
  "published_at": "2026-09-09T20:55:47+00:00",
  "transcript": "*Surveillance plagiarism* - Hosted AI company pumps their stock price by training upon researchers' AI sessions, so that their internal model can solve problems with seemingly less human guidance, but really the model exploits past guidance given by (multiple) humans focused upon problems considered important.\n\nAs background, [Tristan Buckmaster](https://mastodon.social/@tristanbuckmaster/117233413705701198) released a [statement](https://cims.nyu.edu/~tristanb/statement.pdf) about several unethical actions by OpenAI & Sebastian Bubeck, including threats and pushing him to kick his Anthropic coauthor off a paper, but the interesting part for people here:  \n\nAs [clarified by Talia Ringer](https://mastodon.social/@TaliaRinger@mathstodon.xyz/117235246523045723), OpenAI does train upon your uploaded data and your OpenAI sessions, unless you out-out somehow.  This means their internal models could exploit your past prompting work to look more autonomous & intelligent.  \n\nThis is a major confirmation that folks should use locally run open weights models, especially whenever being first or not leaking data matters.  \n\nAll this casts serious doubt upon claim that internal models solved difficult problems largely unaided by humans.  Those hosted AI companies might not even know from where the human prompting originates.\n\n\n\n--- Top Comments ---\n\n\n[43 upvotes] OpenAI stealing IP and data? Color me shocked. \n\nAll of these people building stuff using OAI and Antrophic are basically making sure than 6 months down the line everyone will build the exact same stuff.\n\n[30 upvotes] the people behind claude do the same thing.\n\nnothing new, if you had 5 brain cells, this was well known they would do this\n\n[11 upvotes] One should assume the AI companies are watching everything their users put into frontier models and will take whatever they wish.  This includes prompts in the various walled gardens too, I don't believe for a second that a mere \"privacy\" toggle will stop these sociopaths.  \n\n[10 upvotes] Dont get me wrong, I am all for local models but should the « unless you opt-out somehow » not be considered basic knowledge at that point.\n\nI have not the full details but if as you say they did not opt-out, who should be blamed?",
  "transcript_chars": 2247,
  "ingested_at": "2026-09-10T01:30:03.915398+00:00",
  "source": "reddit",
  "yt_meta": {
    "score": 73,
    "upvote_ratio": 0.85,
    "num_comments": 35,
    "author": "Shoddy-Childhood-511",
    "is_self": true
  }
}