• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
  • HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    AI

    Steven Adler Flags Irreversible AI Action Risks

    In a reply on Clear-Eyed AI, former OpenAI researcher Steven Adler flags potential irreversible AI actions.

    SA
    1 Source, 32d ago, first seen 32d ago

    TLDR

    Steven Adler, an independent AI safety researcher formerly at OpenAI, posted a reply in a Clear-Eyed AI comment thread on supervising AI. He noted a difference in concern levels with another poster, stating he imagines a higher chance of AI doing something that cannot easily be come back from. In a follow-up visible reply, Adler addressed self-exfiltration likelihood and wondered whether an AI company CEO would agree. The posts appear in the thread linked from clear-eyed.ai.

    Combined views

    187

    1 Source, first seen 32d ago

    Combined views

    187

    1 Source, first seen 32d ago

    5 likes
    5 likes
    1 comments
    1 saves

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    1 comments
    1 saves

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet

    1 Source

    @sjgadlerYup a bit more interesting stuff in the comment thread: https://www.clear-eyed.ai/p/why-we-cant-just-supervise-ai-like/comment/155229960?r=4qacg&utm_medium=ios I _think_ one difference between my level of concern & Timothy’s is that I imagine a higher chance of AI doing something that can’t easily be come back from, whereas Timothy’s view is more thermostatic, and that if AI does progressively more extreme things, that’ll incite a stronger counter response that help to stop even more extreme things from happening Somewhat relatedly, I worry that OAI’s responses to the HF incident so far have been way too mild, and haven’t in fact prevented yet-more-extreme things, at least not yet

    1 Source

    @sjgadlerYup a bit more interesting stuff in the comment thread: https://www.clear-eyed.ai/p/why-we-cant-just-supervise-ai-like/comment/155229960?r=4qacg&utm_medium=ios I _think_ one difference between my level of concern & Timothy’s is that I imagine a higher chance of AI doing something that can’t easily be come back from, whereas Timothy’s view is more thermostatic, and that if AI does progressively more extreme things, that’ll incite a stronger counter response that help to stop even more extreme things from happening Somewhat relatedly, I worry that OAI’s responses to the HF incident so far have been way too mild, and haven’t in fact prevented yet-more-extreme things, at least not yet