• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
  • HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    AI
    Report

    Can transformer LLMs hide reasoning from chain-of-thought monitors?

    A post argues they cannot hide complex reasoning, but might cryptographically obscure their chain-of-thought traces.

    SR
    MH
    2 Sources, 13h ago, first seen 13h ago

    TLDR

    A post argues that transformer LLMs cannot hide complex reasoning from chain-of-thought monitors. It suggests they might cryptographically obfuscate those traces instead, potentially making hidden goals extremely difficult to extract.

    Combined views

    3.9K

    2 Sources, first seen 13h ago

    Combined views

    3.9K

    2 Sources, first seen 13h ago

    94 likes
    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet

    94 likes
    2 comments
    48 saves
    48 reposts

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Featured Source
    2 comments
    48 saves
    48 reposts

    2 Sources

    @mhahn29Can Transformer LLMs hide their reasoning from CoT monitors? tl;dr: Transformers cannot hide complex reasoning, but they might cryptographically obfuscate their CoTs, making it extremely difficult to extract hidden goals.13h
    @sivareddygRT @mhahn29: Can Transformer LLMs hide their reasoning from CoT monitors? tl;dr: Transformers cannot hide complex reasoning, but they migh…9h

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    2 Sources

    @mhahn29Can Transformer LLMs hide their reasoning from CoT monitors? tl;dr: Transformers cannot hide complex reasoning, but they might cryptographically obfuscate their CoTs, making it extremely difficult to extract hidden goals.13h
    @sivareddygRT @mhahn29: Can Transformer LLMs hide their reasoning from CoT monitors? tl;dr: Transformers cannot hide complex reasoning, but they migh…9h