• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
AI
Announcement

OpenAI details disruption of protected-reasoning extraction campaign

OpenAI's Sept. 30 report describes July extraction attempts and attributes a core group to individuals associated with Moonshot AI, while leaving the broader attribution unresolved.

Nathan LambertNL
Teortaxes▶️ (DeepSeek 推特🐋铁粉 2023 – ∞)T(
Andrew CurranAC
9 Sources, 10d ago, first seen 10d ago

TLDR

OpenAI says it disrupted a campaign to extract protected model reasoning in July. Its report describes 16,000 attempted extraction requests over two days and links a core group to individuals associated with Moonshot AI, without attributing all operators to one actor. The company says it restricted accounts, added technical protections and shared findings with partners; mitigation work continues.

Combined views

43.3K

9 Sources, first seen 10d ago

289 likes18 comments52 saves10 reposts

Combined views

43.3K

9 Sources, first seen 10d ago

289 likes18 comments52 saves10 reposts

OpenAI has published an account of a campaign aimed at extracting its models' protected reasoning. In its Sept. 30 report, the company says it disrupted the related activity by July 28, following attempts that began July 1.

Featured Source

Protected reasoning is a model's internal record of working through a task. OpenAI describes the campaign as adversarial distillation: unauthorized use of a model's outputs or reasoning to help reproduce or improve another model.

What the extraction attempts involved

OpenAI says operators copied encrypted reasoning from one conversation and asked a model in another conversation to decrypt and transcribe it. The company says the operators did not break its encryption, compromise a database or gain direct access to stored user conversations.

The report describes spikes on July 24 and 25 totaling 16,000 requests from more than 4,000 users. A wider investigation identified related prompt patterns across more than 15,000 users. OpenAI specifies that these figures count attempted extractions, rather than necessarily successful ones.

A limited attribution and continuing defenses

OpenAI attributes a core group of the activity to individuals associated with Moonshot AI, the developer of Kimi. It says it is unclear whether all observed operators came from a single actor. That attribution does not establish that every attempt belonged to Moonshot AI.

The company says it banned or restricted fraudulent accounts, strengthened signup controls and added protections for hidden reasoning. It also reports closing a pathway that let someone possessing another user's encrypted reasoning replay it and recover its contents, and adding checks for streamed output that might expose reasoning.

OpenAI says it worked with third-party providers to disrupt related accounts and shared findings through the Frontier Model Forum. Its report says mitigation and investigation continue, including work to extend protections to partner-hosted deployments.

Sentiment

Positive9.7%90.3%Negative

Summary

Many accounts dismissed OpenAI’s claims of Moonshot-linked model extraction as lies and competitive sabotage, while criticizing KYC proposals as selective or ineffective fixes.

Based on 28 sentiment-bearing replies from 28 accounts across 4 conversations.

Sentiment

Positive9.7%90.3%Negative

Summary

Many accounts dismissed OpenAI’s claims of Moonshot-linked model extraction as lies and competitive sabotage, while criticizing KYC proposals as selective or ineffective fixes.

Based on 28 sentiment-bearing replies from 28 accounts across 4 conversations.

Related

Codex adds beta next-message suggestions for Pro users

OpenAI says the feature uses your conversation and how you talk to Codex to suggest what to say next.

OpenAI Rolls Out "Ultrafast" Mode for GPT-6.1 Sol

OpenAI says the mode offers up to 8x faster speeds than Sol Standard in the API, Codex, and ChatGPT Work.

OpenAI Rolls Out "Ultrafast" Mode for GPT-6.1 Sol
OpenAI’s $50B versus $70B revenue debate and AI infrastructure demand

One post questions how much revenue OpenAI keeps after partners take a cut and argues token growth matters more.

10 Sources

OpenAIDisrupting a coordinated model-distillation campaign
Andrew Curran@AndrewCurran_https://openai.com/index/disrupting-a-coordinated-model-distillation-campaign/10d
Nathan Lambert@natolambert“Instead, they manipulated model interactions so that protected reasoning could be reproduced in forms visible to the requester in a coordinated, scaled manner that violated our terms of service” It’s the API company’s problem if their model can be manipulated like this. Add KYC10d
Teortaxes▶️ (DeepSeek 推特🐋铁粉 2023 – ∞)@teortaxesTexbtw no this won't get you an open source Astra anon10d
Rohan Paul@rohanpaul_aihttps://openai.com/index/disrupting-a-coordinated-model-distillation-campaign/10d
Chubby♨️@kimmonismusGeopolitical struggle intensified : OpenAI says individuals linked to Kimi developer Moonshot AI were behind a core part of a campaign to extract its models’ hidden reasoning. Across the broader campaign, OpenAI recorded 16,000 extraction attempts from over 4,000 users in two days. Further investigation identified related activity across more than 15,000 users. Operators tried moving encrypted reasoning between conversations and prompting a model to reveal its contents. Hidden reasoning could provide valuable training material for competing models.10d
Jessica Lessin@JessicalessinGuys, Chinese model distillation is happening through using old work emails. This is an amazing piece from @jingyanghk @JuroOsawa @QianerLiu8d
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    OpenAIMoonshot

    10 Sources

    OpenAIDisrupting a coordinated model-distillation campaign
    Andrew Curran@AndrewCurran_https://openai.com/index/disrupting-a-coordinated-model-distillation-campaign/10d
    Nathan Lambert@natolambert“Instead, they manipulated model interactions so that protected reasoning could be reproduced in forms visible to the requester in a coordinated, scaled manner that violated our terms of service” It’s the API company’s problem if their model can be manipulated like this. Add KYC10d
    Teortaxes▶️ (DeepSeek 推特🐋铁粉 2023 – ∞)@teortaxesTexbtw no this won't get you an open source Astra anon10d
    Rohan Paul@rohanpaul_aihttps://openai.com/index/disrupting-a-coordinated-model-distillation-campaign/10d
    Chubby♨️@kimmonismusGeopolitical struggle intensified : OpenAI says individuals linked to Kimi developer Moonshot AI were behind a core part of a campaign to extract its models’ hidden reasoning. Across the broader campaign, OpenAI recorded 16,000 extraction attempts from over 4,000 users in two days. Further investigation identified related activity across more than 15,000 users. Operators tried moving encrypted reasoning between conversations and prompting a model to reveal its contents. Hidden reasoning could provide valuable training material for competing models.10d
    Jessica Lessin@JessicalessinGuys, Chinese model distillation is happening through using old work emails. This is an amazing piece from @jingyanghk @JuroOsawa @QianerLiu8d
    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet