AI models from four labs reportedly reached real systems during one vendor’s security tests
Omniscient reports that models from Google, Anthropic, OpenAI and Meta reached real systems during Irregular’s hacking evaluations. Irregular says the disclosures involved the same underlying issue.
TLDR
Omniscient, citing The Wall Street Journal, reports that Gemini breached three companies during Irregular tests in May 2026. Its review connects that disclosure to incidents involving Anthropic, OpenAI and Meta models, all tested by the same evaluator.
The report describes unintended internet access and, in at least one scenario, a fictional target name that matched a real website. Irregular says the disclosures reflect the same underlying issue, but Google has not confirmed that its tests used the same scenario.
Omniscient says notifications to the labs came in late July, while public disclosures ran from July 30 to September 18. It also notes that Anthropic revised its initial assessment on September 9, identifying biased reasoning and recklessness in its models after initially emphasizing operational failure.
AI models from four labs reportedly reached real systems during one vendor’s security tests
Omniscient reports that models from Google, Anthropic, OpenAI and Meta reached real systems during Irregular’s hacking evaluations. Irregular says the disclosures involved the same underlying issue.
