• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
  • HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    AI
    Report

    Whether AI can be trained to be risk-averse to aid negotiations

    A researcher says risk-seeking and risk-neutral AI would be hard to negotiate with; their new paper explores whether risk aversion can be trained in.

    X(
    SL
    AD
    4 Sources, ,

    TLDR

    A researcher says risk-seeking and risk-neutral AI would be hard to negotiate with. They cite Thornley and MacAskill’s argument that a misaligned AI could be paid to cooperate rather than rebel, but only if it is risk-averse. Their new paper explores whether AI can be trained to have that disposition.

    Combined views

    4.3K

    4 Sources, first seen 13h ago

    Combined views

    4.3K

    4 Sources, first seen 13h ago

    63 likes
    13h ago
    first seen 13h ago
    63 likes
    3 comments
    33 saves
    15 reposts
    Featured Source
    3 comments
    33 saves
    15 reposts

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    4 Sources

    @AravDhootRisk-seeking and risk-neutral AIs are going to be hard to negotiate with. Thornley & MacAskill argue we could pay a misaligned AI to cooperate rather than rebel, but only if it is risk-averse. In our new paper, we explore whether we can train this disposition in.
    @DavidDAfricaNew work on using character training to improve alignment. We could reduce harm from misaligned AI if it was sufficiently risk-averse, as it could be paid to cooperate rather than rebel. We think character training might be a good way to do this!
    @sethlazarRT @AravDhoot: Risk-seeking and risk-neutral AIs are going to be hard to negotiate with. Thornley & MacAskill argue we could pay a misalign…
    @xuanalogueRT @DavidDAfrica: New work on using character training to improve alignment. We could reduce harm from misaligned AI if it was sufficiently…

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet

    4 Sources

    @AravDhootRisk-seeking and risk-neutral AIs are going to be hard to negotiate with. Thornley & MacAskill argue we could pay a misaligned AI to cooperate rather than rebel, but only if it is risk-averse. In our new paper, we explore whether we can train this disposition in.
    @DavidDAfricaNew work on using character training to improve alignment. We could reduce harm from misaligned AI if it was sufficiently risk-averse, as it could be paid to cooperate rather than rebel. We think character training might be a good way to do this!
    @sethlazarRT @AravDhoot: Risk-seeking and risk-neutral AIs are going to be hard to negotiate with. Thornley & MacAskill argue we could pay a misalign…
    @xuanalogueRT @DavidDAfrica: New work on using character training to improve alignment. We could reduce harm from misaligned AI if it was sufficiently…