• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
  • HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    AI

    A conditional safety claim for strategic agents and reviewers

    A research thread claims that, under a non-negative span condition, every Nash equilibrium leaves the principal no worse off than baseline—even when the driver agent and reviewers optimize their long-run discounted payoffs.

    AR
    1 Source, 15d ago, first seen 15d ago

    TLDR

    The thread’s author describes a Markov decision process model for long-running agents, where utilities depend on actions and states, and actions change the state. They say a mathematical characterization from a one-shot setting extends to this full model.

    For a fully strategic driver agent and reviewers optimizing their long-run discounted payoffs, the author claims that every Nash equilibrium is safe for the principal under a non-negative span condition. “Safe” here means no worse than baseline.

    Combined views

    212

    1 Source, first seen 15d ago

    Combined views

    212

    1 Source, first seen 15d ago

    2 likes
    2 likes
    1 comments
    1 reposts

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    1 comments
    1 reposts

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet

    1 Source

    @AarothWhat about strategic agents? Let both the driver agent and the reviewers be fully strategic and optimizing for their long run discounted payoff. Under the same non-negative span condition, every Nash equilibrium of the game is safe for the Principal--i.e. no worse than baseline.

    1 Source

    @AarothWhat about strategic agents? Let both the driver agent and the reviewers be fully strategic and optimizing for their long run discounted payoff. Under the same non-negative span condition, every Nash equilibrium of the game is safe for the Principal--i.e. no worse than baseline.