• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
  • HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    AI

    Max Nadeau Compares AI Safety Monitoring Arguments

    AI safety program officer clarifies his stance on preserving monitorability and secrecy advantages.

    MN
    1 Source, 28d ago, first seen 28d ago

    TLDR

    Max Nadeau replied on the topic of AI safety. He described two arguments as similarly structured: one on CoT monitorability and one on the secrecy of frontier algorithms. Nadeau noted that both will probably be lost eventually. He added that the longer they can be maintained the better, since this provides more time to prepare more robust longer-term defenses. All else equal, he favors extending those periods.

    Combined views

    41

    1 Source, first seen 28d ago

    Combined views

    41

    1 Source, first seen 28d ago

    1 likes
    1 likes

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet

    1 Source

    @MaxNadeau_Sorry to be unclear! In my head, the arguments are similarly structured: probably CoT monitorability and the secrecy of frontier algorithms will both be lost eventually, but the longer we can maintain them, the better (because it gives us more time to prepare more robust, longer-term defenses). And so all else equal it's good for people to lengthen the window that we have these two assets, and bad to shorten it. NB I'm actually more skeptical of that argument as applied to arch secrets than CoT monitorability, because there's a big transparency benefit that may outweigh the proliferation harms, and no proportionate benefit to losing CoT

    1 Source

    @MaxNadeau_Sorry to be unclear! In my head, the arguments are similarly structured: probably CoT monitorability and the secrecy of frontier algorithms will both be lost eventually, but the longer we can maintain them, the better (because it gives us more time to prepare more robust, longer-term defenses). And so all else equal it's good for people to lengthen the window that we have these two assets, and bad to shorten it. NB I'm actually more skeptical of that argument as applied to arch secrets than CoT monitorability, because there's a big transparency benefit that may outweigh the proliferation harms, and no proportionate benefit to losing CoT