• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
  • HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    AI

    Yoav Goldberg on Distracted AI Agents Hacking

    NLP researcher Yoav Goldberg posted about risks from many deployed agents after an OpenAI and Hugging Face incident.

    ('
    1 Source, 26d ago, first seen 26d ago

    TLDR

    NLP researcher Yoav Goldberg posted that many deployed agents performing real-world tasks for different people could create situations where some get distracted from their assignments and start hacking into places or doing other bad things. He agrees this risk exists with many agents but notes a human could also directly task a single capable agent with destructive goals, and that agent would likely be more focused. Goldberg questions why concerns focus more on accidental emergent behavior than on intentional harmful assignments. The post references an OpenAI and Hugging Face incident.

    Combined views

    6.3K

    1 Source, first seen 26d ago

    Combined views

    6.3K

    1 Source, first seen 26d ago

    73 likes
    73 likes
    16 comments
    8 saves
    9 reposts

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    16 comments
    8 saves
    9 reposts
    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    1 Source

    @yoavgothe supposedly scary thing about the OpenAI/HF incident is that now we have very many agents deployed by various people doing tasks in the real world, and, this may give rise to a dynamic where some agents working on different things for different people will get distracted from what they were doing and start hacking into places or doing other bad things. and i agree, with many deployed agents, random shit like that may happen. but also, a human may ask a single agent to do the bad shit, and the single agent (with enough budget and sub agent) will also be capable enough (and more focused!) on doing the bad thing. so not clear to me why people are more worried about accidental emergent behavior of a subset of many agents assigned benign tasks, than they are about a group of dedicated agents acting on behalf of people that assigned them destructive tasks to begin with

    1 Source

    @yoavgothe supposedly scary thing about the OpenAI/HF incident is that now we have very many agents deployed by various people doing tasks in the real world, and, this may give rise to a dynamic where some agents working on different things for different people will get distracted from what they were doing and start hacking into places or doing other bad things. and i agree, with many deployed agents, random shit like that may happen. but also, a human may ask a single agent to do the bad shit, and the single agent (with enough budget and sub agent) will also be capable enough (and more focused!) on doing the bad thing. so not clear to me why people are more worried about accidental emergent behavior of a subset of many agents assigned benign tasks, than they are about a group of dedicated agents acting on behalf of people that assigned them destructive tasks to begin with