• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
  • HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    AI

    Rumors Claim Chinese Labs Distilled Reasoning Traces

    Tweets discuss speculative claims that Chinese labs distilled open models using hidden reasoning traces.

    JM
    T(
    WB
    4 Sources, 52d ago, first seen 52d ago

    TLDR

    Jack Morris posted about rumors that Chinese labs like those behind Kimi, Qwen, and Minimax have extracted reasoning traces from open-weights models for better distillation. He tied the claims to his own PhD research and noted the traces are normally hidden from users. Teortaxes replied that the process resembles stone soup distillation rather than meaningful chain-of-thought exfiltration. The exchange also references an arXiv paper titled How to Steal Reasoning Without Reasoning Traces. No independent confirmation of the rumors appears in the packet.

    Combined views

    116.9K

    4 Sources, first seen 52d ago

    Combined views

    116.9K

    4 Sources, first seen 52d ago

    1.4K likes
    1.4K likes
    49 comments
    1.3K saves
    108 reposts

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Featured Source
    49 comments
    1.3K saves
    108 reposts

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet

    4 Sources

    @jxmnopthere are some really interesting rumors going around related to the distillation of open-weights models (Kimi, Qwen, Minimax, etc.) and they're very related to my PhD work The narrative [speculative]: • good distillation relies on reasoning traces, normally hidden from users • Chinese labs figured out in early 2026 how to reverse-engineer reasoning from Claude Code and Codex • they were able to collect large amounts of long-horizon data *with reasoning traces included* this way • this jailbreak led to a new wave of OSS models we've enjoyed over the past few months I'm not sure how true it is, but reasoning extractability seems like a huge uncertainty around the future of open models in particular i'm curious how much having the reasoning matters. this (plus Anthropic's messaging around distillation attacks) indicates that reasoning chains are crucial for distilling model capabilities. our research (http://arxiv.org/abs/2603.07267) found something different: if you train a high-quality reasoning inverter, it's often pretty easy to reconstruct useful traces from frontier models given their outputs. figuring out how to approximate frontier model reasoning traces might turn out to be an existential problem for open weights models
    @teortaxesTexIsn't this stone soup distillation? At this point, just synthesize ground truth as well, my dog. This is not CoT exfiltration in a sense that matters. What are we doing here
    @willcb@teortaxesTex VR-CLI is the golden path, with maybe a little bit of this stuff as a self-distill nudge to get off the ground

    4 Sources

    @jxmnopthere are some really interesting rumors going around related to the distillation of open-weights models (Kimi, Qwen, Minimax, etc.) and they're very related to my PhD work The narrative [speculative]: • good distillation relies on reasoning traces, normally hidden from users • Chinese labs figured out in early 2026 how to reverse-engineer reasoning from Claude Code and Codex • they were able to collect large amounts of long-horizon data *with reasoning traces included* this way • this jailbreak led to a new wave of OSS models we've enjoyed over the past few months I'm not sure how true it is, but reasoning extractability seems like a huge uncertainty around the future of open models in particular i'm curious how much having the reasoning matters. this (plus Anthropic's messaging around distillation attacks) indicates that reasoning chains are crucial for distilling model capabilities. our research (http://arxiv.org/abs/2603.07267) found something different: if you train a high-quality reasoning inverter, it's often pretty easy to reconstruct useful traces from frontier models given their outputs. figuring out how to approximate frontier model reasoning traces might turn out to be an existential problem for open weights models
    @teortaxesTexIsn't this stone soup distillation? At this point, just synthesize ground truth as well, my dog. This is not CoT exfiltration in a sense that matters. What are we doing here
    @willcb@teortaxesTex VR-CLI is the golden path, with maybe a little bit of this stuff as a self-distill nudge to get off the ground