• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
  • HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    AI

    Better oversight methods versus more evaluator organizations for frontier AI

    The post invites collaborators to develop oversight that can scale, including ways to verify AI agents’ behavior and study collusion and covert communication in systems with multiple agents.

    TG
    VD
    4 Sources, ,

    TLDR

    A call for collaborators argues that adding evaluator organizations is not enough for effective frontier AI oversight. Topics of interest include formal and runtime verification, statistical guarantees for oversight, and supervision of very large groups of AI agents. The post also highlights collusion, covert communication and emerging cooperative norms among agents, alongside alternatives to monitoring an AI’s chain of thought and methods for AI control. It links to a form for people interested in working together.

    Combined views

    6K

    4 Sources, first seen 18d ago

    Combined views

    6K

    4 Sources, first seen 18d ago

    39 likes
    18d ago
    first seen 18d ago
    39 likes
    6 comments
    40 saves
    5 reposts

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    6 comments
    40 saves
    5 reposts

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet

    4 Sources

    @timrudnerWe don't just need more evaluator organizations, we also need better methods for effective frontier AI oversight. Fill out this form if you want to work together on improving scalable oversight for frontier AI: https://forms.gle/VpWyH6gTqqcMot7e9 Topics of interest include (but are not limited to): - Formal verification and runtime verification for AI agents and multi-agent systems - Scalable oversight for very large multi-agent systems - Statistical guarantees for scalable oversight - Mechanism design for frontier AI multi-agent systems - Collusion in frontier AI multi-agent systems - Emergent cooperative norms in frontier AI multi-agent systems - Covert communication (steganography) in frontier AI multi-agent systems - Oversight alternatives to CoT monitoring - AI control
    @vasishtdudduWe are looking for people interested to work on AI safety!! Please see more details below

    4 Sources

    @timrudnerWe don't just need more evaluator organizations, we also need better methods for effective frontier AI oversight. Fill out this form if you want to work together on improving scalable oversight for frontier AI: https://forms.gle/VpWyH6gTqqcMot7e9 Topics of interest include (but are not limited to): - Formal verification and runtime verification for AI agents and multi-agent systems - Scalable oversight for very large multi-agent systems - Statistical guarantees for scalable oversight - Mechanism design for frontier AI multi-agent systems - Collusion in frontier AI multi-agent systems - Emergent cooperative norms in frontier AI multi-agent systems - Covert communication (steganography) in frontier AI multi-agent systems - Oversight alternatives to CoT monitoring - AI control
    @vasishtdudduWe are looking for people interested to work on AI safety!! Please see more details below