• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
  • HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    AI

    The evidence behind claims of AI deception and shutdown resistance

    A coauthor of an ICML position paper argues that current evidence often can't distinguish apparent deception or shutdown resistance from role-play, instruction-following or task-completion pressure.

    GM
    1 Source, 15d ago, first seen 15d ago

    TLDR

    A coauthor says the position paper maps where evidence gets thin across four stages and proposes a shared standard to strengthen it. The concern is that claims about AI deception and shutdown resistance are starting to inform deployment and regulation, even though evidence often can't separate those interpretations from role-play, instruction-following or task-completion pressure. On June 18, 2026, the coauthor announced that the paper had been accepted for an oral presentation at ICML.

    Combined views

    8.6K

    1 Source, first seen 15d ago

    Combined views

    8.6K

    1 Source, first seen 15d ago

    64 likes
    64 likes
    6 comments
    30 saves
    17 reposts

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    6 comments
    30 saves
    17 reposts

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet

    1 Source

    @GaryMarcus“When a model looks like it "deceives" or "resists shutdown," how do we know it isn't role-play, instruction-following, or just task-completion pressure? Often the current evidence can't yet tell them apart, and those claims are starting to inform deployment and regulation.” - @XinCynthiaChen Excellent thread, with important recommendations for researchers in how they talk about their work.

    1 Source

    @GaryMarcus“When a model looks like it "deceives" or "resists shutdown," how do we know it isn't role-play, instruction-following, or just task-completion pressure? Often the current evidence can't yet tell them apart, and those claims are starting to inform deployment and regulation.” - @XinCynthiaChen Excellent thread, with important recommendations for researchers in how they talk about their work.