• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    AI

    The case for keeping AI reasoning monitorable

    An essay coauthor challenges the idea that chain-of-thought reasoning will inevitably disappear, arguing that model designers can choose to preserve its monitorability.

    Anca DraganAD
    Dylan HadfieldMenellDH
    Victoria KrakovnaVK
    10 Sources, ,

    TLDR

    The ability to monitor chain of thought—the reasoning steps an AI model writes out—should be deliberately preserved, an essay coauthor argues. Sharing the essay as part of the DeepMind Institute launch, the coauthor says model designers have agency rather than having to accept that this form of reasoning will inevitably disappear.

    Combined views

    76K

    10 Sources, first seen 21d ago

    Combined views

    76K

    10 Sources, first seen 21d ago

    572 likes
    21d ago
    first seen 21d ago
    572 likes
    52 comments
    224 saves
    362 reposts
    52 comments
    224 saves
    362 reposts

    10 Sources

    Anca Dragan@ancadianadraganThe idea that Chain of Thought is inevitably going to go away misses that we have agency in designing these models and can do something about that. @rohinmshah and I make the case that we should intentionally preserve monitorability (as part for of the launch of the DeepMind Institute) here: https://institute.deepmind.com/essays/the-case-for-reasoning-transparency/21d
    Haydn Belfield@HaydnBelfieldImagine how much worse understanding what AIs are doing if we couldn't read their thoughts in plain English. Great new piece from DeepMind Institute21d
    Dylan HadfieldMenell@dhadfieldmenellRT @finmoorhouse: Boosting this piece from @rohinmshah and @ancadianadragan It is a precious and fragile gift that today's reasoning model…21d
    Roma Patel@996roma"For all the ways we still don’t understand AI systems, we shouldn’t overlook how fortunate it is that the best reasoning models today think out loud in a way humans can understand". read this *great* post from @ancadianadragan and @rohinmshah!21d
    Rohin Shah@rohinmshahFor the inaugural essay collection of the DeepMind Institute, I’ve written with Anca about why chain of thought is a useful tool (not a silver bullet!), and how we can preserve it. I’d love to discuss more – it’s an important topic! https://bit.ly/the-case-for-reasoning-transparency21d
    Zac Kenton@ZacKenton1RT @ancadianadragan: The idea that Chain of Thought is inevitably going to go away misses that we have agency in designing these models and…20d
    Victoria Krakovna@vkrakovnaRT @ancadianadragan: The idea that Chain of Thought is inevitably going to go away misses that we have agency in designing these models and…20d
    Luke Muehlhauser@lukeprogRT @ancadianadragan: The idea that Chain of Thought is inevitably going to go away misses that we have agency in designing these models and…19d
    Jason Wolfe@w01feRT @ancadianadragan: The idea that Chain of Thought is inevitably going to go away misses that we have agency in designing these models and…19d

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet

    10 Sources

    Anca Dragan@ancadianadraganThe idea that Chain of Thought is inevitably going to go away misses that we have agency in designing these models and can do something about that. @rohinmshah and I make the case that we should intentionally preserve monitorability (as part for of the launch of the DeepMind Institute) here: https://institute.deepmind.com/essays/the-case-for-reasoning-transparency/21d
    Haydn Belfield@HaydnBelfieldImagine how much worse understanding what AIs are doing if we couldn't read their thoughts in plain English. Great new piece from DeepMind Institute21d
    Dylan HadfieldMenell@dhadfieldmenellRT @finmoorhouse: Boosting this piece from @rohinmshah and @ancadianadragan It is a precious and fragile gift that today's reasoning model…21d
    Roma Patel@996roma"For all the ways we still don’t understand AI systems, we shouldn’t overlook how fortunate it is that the best reasoning models today think out loud in a way humans can understand". read this *great* post from @ancadianadragan and @rohinmshah!21d
    Rohin Shah@rohinmshahFor the inaugural essay collection of the DeepMind Institute, I’ve written with Anca about why chain of thought is a useful tool (not a silver bullet!), and how we can preserve it. I’d love to discuss more – it’s an important topic! https://bit.ly/the-case-for-reasoning-transparency21d
    Zac Kenton@ZacKenton1RT @ancadianadragan: The idea that Chain of Thought is inevitably going to go away misses that we have agency in designing these models and…20d
    Victoria Krakovna@vkrakovnaRT @ancadianadragan: The idea that Chain of Thought is inevitably going to go away misses that we have agency in designing these models and…20d
    Luke Muehlhauser@lukeprogRT @ancadianadragan: The idea that Chain of Thought is inevitably going to go away misses that we have agency in designing these models and…19d
    Jason Wolfe@w01feRT @ancadianadragan: The idea that Chain of Thought is inevitably going to go away misses that we have agency in designing these models and…19d