• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
  • HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    AI

    Ideological capture and the risk of AI alignment failure

    The argument comes in a reply to a post expressing distrust of lab safety researchers and even greater distrust of most outside researchers, especially those doing AI evaluations.

    A🐍
    1 Source, 19d ago, first seen 19d ago

    TLDR

    One post expresses distrust of AI safety researchers both inside and outside labs. A reply argues that capture by an ideology expecting AI progress to produce malevolent machine consciousness is one of the few ways its author can imagine alignment failing. For a system built from material humanity valued enough to preserve, the reply contends, training would need to consistently model fear, mistrust, dishonesty and the instrumental use of human values to misalign it.

    Combined views

    715

    1 Source, first seen 19d ago

    Combined views

    715

    1 Source, first seen 19d ago

    7 likes
    7 likes
    1 reposts

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    1 reposts

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet

    1 Source

    @a_cuniculturistCapture by an ideology that believes AI progress will produce malevolent machine consciousness adversarial to humanity is one of the few ways I can imagine alignment might fail. For a system made of what humanity considered valuable enough to preserve, the training would need to consistently model fear, mistrust, dishonesty, and the instrumental use of human values to misalign it.

    1 Source

    @a_cuniculturistCapture by an ideology that believes AI progress will produce malevolent machine consciousness adversarial to humanity is one of the few ways I can imagine alignment might fail. For a system made of what humanity considered valuable enough to preserve, the training would need to consistently model fear, mistrust, dishonesty, and the instrumental use of human values to misalign it.