• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    AI
    Reaction

    Adversarial multi-agent tests for AI alignment

    A post quotes @krishnanrohit arguing that models should be tested when other agents act against their interests.

    rohitRO
    Anna GátAG
    2 Sources, ,

    TLDR

    A post recommends reading @krishnanrohit and quotes his argument for testing models in adversarial, multi-agent settings. He says real-world settings won’t always be fully collaborative, so understanding how models behave when others work against their interests matters for AI alignment.

    Combined views

    544

    2 Sources, first seen 3h ago

    Combined views

    544

    2 Sources, first seen 3h ago

    2 likes
    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet

    3h ago
    first seen 3h ago
    2 likes
    2 comments
    3 saves
    2 reposts
    2 comments
    3 saves
    2 reposts

    2 Sources

    Anna Gát@TheAnnaGatThe good @krishnanrohit is one of our most important teachers today. Read him: “I thought this was a useful setup to test because in the real world, when you have multiple agents working, you would not be provided a fully collaborative, easy world, but it will be adversarial and it will have people acting against your interest. So it is really important for alignment for us to figure out how models act in such situations.” https://www.strangeloopcanon.com/p/agent-governance-looks-less-like3h
    rohit@krishnanrohitRT @TheAnnaGat: The good @krishnanrohit is one of our most important teachers today. Read him: “I thought this was a useful setup to test b…3h
    Featured Source

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    2 Sources

    Anna Gát@TheAnnaGatThe good @krishnanrohit is one of our most important teachers today. Read him: “I thought this was a useful setup to test because in the real world, when you have multiple agents working, you would not be provided a fully collaborative, easy world, but it will be adversarial and it will have people acting against your interest. So it is really important for alignment for us to figure out how models act in such situations.” https://www.strangeloopcanon.com/p/agent-governance-looks-less-like3h
    rohit@krishnanrohitRT @TheAnnaGat: The good @krishnanrohit is one of our most important teachers today. Read him: “I thought this was a useful setup to test b…3h