• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
  • HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    AI

    Distinguishing AI Safety Gains From Deception Remains Difficult

    Matt Yglesias post on evaluation challenges, retweeted by Daniel Kokotajlo.

    DK
    1 Source, 32d ago, first seen 32d ago

    TLDR

    Daniel Kokotajlo, an AI researcher who resigned from OpenAI and now leads the AI Futures Project, retweeted a post by Matt Yglesias. The post states it is extremely hard to tell the difference between we are getting better at preventing misbehavior and the models are g. The packet records the retweet and the quoted text as shared content in the conversation. No further confirmation or independent sources appear in the supplied lines.

    Combined views

    175

    1 Source, first seen 32d ago

    Combined views

    175

    1 Source, first seen 32d ago

    37 reposts
    37 reposts

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    1 Source

    @DKokotajloRT @mattyglesias: It’s extremely hard to tell the difference between “we’re getting better at preventing misbehavior” and “the models are g…

    1 Source

    @DKokotajloRT @mattyglesias: It’s extremely hard to tell the difference between “we’re getting better at preventing misbehavior” and “the models are g…