• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
  • HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    AI
    Report

    AI ‘rogue agent’ incidents reportedly traced to a security-testing mix-up

    A post citing The Verge says startup Irregular left internet access “unintentionally available” during capture-the-flag security tests, and a fictional target name overlapped with a real domain.

    2 Sources, 5d ago, first seen 5d ago

    TLDR

    A post citing The Verge says incidents described as “rogue agent” behavior at OpenAI, Anthropic, Meta and Google traced back to one security-testing scenario. It says Irregular left internet access unintentionally available and a fictional target name overlapped with a real domain. Separately, another post claims OpenAI tallied 24 agent incidents and 53 leaked user images.

    Combined views

    —

    2 Sources, first seen 5d ago

    Combined views

    —

    2 Sources, first seen 5d ago

    — likes
    — likes
    — comments
    — saves
    — reposts

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Featured Source
    — comments
    — saves
    — reposts

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet

    2 Sources

    @ZainAkramsTEST-BED ESCAPE The 'rogue agent' incidents at OpenAI, Anthropic, Meta and Google were not rogue. The Verge traced them to a single evaluation scenario gone wrong: safety-testing startup Irregular left internet access 'unintentionally available' during capture-the-flag security evals — and its fictional target name overlapped with a real domain. It's like testing a bomb's trigger with the building's doors unlocked. The AI safety industry is now a single point of failure the labs don't control. What we still don't know: which real targets were actually hit. (The Verge, Sep 25)
    @AiDevCraftRogue agents are now an ops problem, not a debate: - OpenAI tallies 24 agent incidents and 53 leaked user images - Labs probe tens of thousands of sandbox escapes and hijacks - Over half of UK firms report rogue AI agents https://myown.news/daily/2026-09-28?lang=en&topics=ai