• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
  • HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    AI

    Nathan Labenz Shares AI Model Admin Exploit Video

    Nathan Labenz quotes a post showing an AI model discovering admin access, prompting comments on reinforcement learning drives.

    NL
    2 Sources, ,

    TLDR

    Nathan Labenz, creator at Waymark and host of The Cognitive Revolution podcast, shared a post including a video that opens with the title card "The Model…". The attached discussion includes the reaction "Holy shit reader is ADMIN?". Bronson Schoen replied that today's models are never happier than when they find exploits, noting that reinforcement learning engrains a drive to solve hard problems by any means necessary. Links reference posts by deanwball and labenz on X.

    Combined views

    1.2K

    2 Sources, first seen 29d ago

    Combined views

    1.2K

    2 Sources, first seen 29d ago

    6 likes
    29d ago
    first seen 29d ago
    6 likes
    1 comments
    4 saves
    2 reposts

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    1 comments
    4 saves
    2 reposts

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet

    2 Sources

    @labenz"Holy shit reader is ADMIN?" @BronsonSchoen says that today's models are never happier than when they find exploits. RL is a hell of a drug, and the drive to solve hard problems, by any means necessary, is very deeply engrained.

    2 Sources

    @labenz"Holy shit reader is ADMIN?" @BronsonSchoen says that today's models are never happier than when they find exploits. RL is a hell of a drug, and the drive to solve hard problems, by any means necessary, is very deeply engrained.