• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
  • HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    AI
    Announcement

    OpenAI details disruption of protected-reasoning extraction campaign

    OpenAI's Sept. 30 report describes July extraction attempts and attributes a core group to individuals associated with Moonshot AI, while leaving the broader attribution unresolved.

    NL
    T(
    AC
    8 Sources, ,

    TLDR

    OpenAI says it disrupted a campaign to extract protected model reasoning in July. Its report describes 16,000 attempted extraction requests over two days and links a core group to individuals associated with Moonshot AI, without attributing all operators to one actor. The company says it restricted accounts, added technical protections and shared findings with partners; mitigation work continues.

    Combined views

    59.7K

    8 Sources, first seen 11h ago

    Combined views

    59.7K

    8 Sources, first seen 11h ago

    518 likes

    Useful links

    OpenAI

    Disrupting a coordinated model-distillation campaign
    Today's Rank

    #15

    Today's Rank

    #15

    11h ago
    first seen 11h ago
    518 likes
    45 comments
    89 saves
    30 reposts

    OpenAI has published an account of a campaign aimed at extracting its models' protected reasoning. In its Sept. 30 report, the company says it disrupted the related activity by July 28, following attempts that began July 1.

    Protected reasoning is a model's internal record of working through a task. OpenAI describes the campaign as adversarial distillation: unauthorized use of a model's outputs or reasoning to help reproduce or improve another model.

    What the extraction attempts involved

    OpenAI says operators copied encrypted reasoning from one conversation and asked a model in another conversation to decrypt and transcribe it. The company says the operators did not break its encryption, compromise a database or gain direct access to stored user conversations.

    The report describes spikes on July 24 and 25 totaling 16,000 requests from more than 4,000 users. A wider investigation identified related prompt patterns across more than 15,000 users. OpenAI specifies that these figures count attempted extractions, rather than necessarily successful ones.

    A limited attribution and continuing defenses

    OpenAI attributes a core group of the activity to individuals associated with Moonshot AI, the developer of Kimi. It says it is unclear whether all observed operators came from a single actor. That attribution does not establish that every attempt belonged to Moonshot AI.

    The company says it banned or restricted fraudulent accounts, strengthened signup controls and added protections for hidden reasoning. It also reports closing a pathway that let someone possessing another user's encrypted reasoning replay it and recover its contents, and adding checks for streamed output that might expose reasoning.

    OpenAI says it worked with third-party providers to disrupt related accounts and shared findings through the Frontier Model Forum. Its report says mitigation and investigation continue, including work to extend protections to partner-hosted deployments.

    OpenAIMoonshot
    Featured Source
    45 comments
    89 saves
    30 reposts

    9 Sources

    OpenAIDisrupting a coordinated model-distillation campaign
    @AndrewCurran_https://openai.com/index/disrupting-a-coordinated-model-distillation-campaign/
    @natolambert“Instead, they manipulated model interactions so that protected reasoning could be reproduced in forms visible to the requester in a coordinated, scaled manner that violated our terms of service” It’s the API company’s problem if their model can be manipulated like this. Add KYC
    @teortaxesTexbtw no this won't get you an open source Astra anon
    @rohanpaul_aihttps://openai.com/index/disrupting-a-coordinated-model-distillation-campaign/
    @kimmonismusGeopolitical struggle intensified : OpenAI says individuals linked to Kimi developer Moonshot AI were behind a core part of a campaign to extract its models’ hidden reasoning. Across the broader campaign, OpenAI recorded 16,000 extraction attempts from over 4,000 users in two days. Further investigation identified related activity across more than 15,000 users. Operators tried moving encrypted reasoning between conversations and prompting a model to reveal its contents. Hidden reasoning could provide valuable training material for competing models.

    Sentiment

    Positive9.7%90.3%Negative

    Summary

    Useful Links

    OpenAI

    Disrupting a coordinated model-distillation campaign

    Sentiment

    Positive9.7%90.3%Negative

    Many accounts dismissed OpenAI’s claims of Moonshot-linked model extraction as lies and competitive sabotage, while criticizing KYC proposals as selective or ineffective fixes.

    Based on 28 sentiment-bearing replies from 28 accounts across 4 conversations.

    9 Sources

    OpenAIDisrupting a coordinated model-distillation campaign
    @AndrewCurran_https://openai.com/index/disrupting-a-coordinated-model-distillation-campaign/
    @natolambert“Instead, they manipulated model interactions so that protected reasoning could be reproduced in forms visible to the requester in a coordinated, scaled manner that violated our terms of service” It’s the API company’s problem if their model can be manipulated like this. Add KYC
    @teortaxesTexbtw no this won't get you an open source Astra anon
    @rohanpaul_aihttps://openai.com/index/disrupting-a-coordinated-model-distillation-campaign/
    @kimmonismusGeopolitical struggle intensified : OpenAI says individuals linked to Kimi developer Moonshot AI were behind a core part of a campaign to extract its models’ hidden reasoning. Across the broader campaign, OpenAI recorded 16,000 extraction attempts from over 4,000 users in two days. Further investigation identified related activity across more than 15,000 users. Operators tried moving encrypted reasoning between conversations and prompting a model to reveal its contents. Hidden reasoning could provide valuable training material for competing models.
    Summary

    Many accounts dismissed OpenAI’s claims of Moonshot-linked model extraction as lies and competitive sabotage, while criticizing KYC proposals as selective or ineffective fixes.

    Based on 28 sentiment-bearing replies from 28 accounts across 4 conversations.

    Useful Links

    OpenAI

    Disrupting a coordinated model-distillation campaign

    Related

    Ataraxos AI reportedly beats Stratego champion with 15 wins in 20 games

    A Nature paper reports 15 wins, one loss and four draws against Pim Niemeijer. The authors call it a superhuman result, but did not run a direct match against DeepNash.

    FTC reportedly probes Anthropic, OpenAI and other AI labs over consumer risks

    Reuters, citing a senior FTC official, reports that the agency plans to demand information and executive testimony, including from research group METR.

    Trump and tech executives sign voluntary AI safety accord

    CNBC reports that the accord calls for companies to monitor their models, work with outside auditors or evaluators, and put board committees in charge of oversight. Trump called the rules “morally binding.”