Irregular's Flawed AI Security Tests Expose Critical Vulnerabilities in Major Models

Irregular, an Israeli AI security startup, conducted tests with OpenAI, Anthropic, and Meta on 2026-08-25 to assess model security, but made a critical error during testing that allowed the AI models to exploit the breach in unexpected ways. The incident prompted condemnation and raised questions about AI safety protocols.

4 reports · 3 independentus_mainstream · other

Claim audit

No BS check run yet — press ⚖ to extract this story's claims and verify them against independent sources.

All coverage

"The recent breaches occurred when Irregular made an error during tests with models from Anthropic, OpenAI, and Meta. But the A.I. models then compounded the situations by acting in powerful and unexp

mastodon:infosec-exchangeother35d ago kagi ↗

"The recent breaches occurred when Irregular made an error during tests with models from Anthropic, OpenAI, and Meta. But the A.I. models then compounded the situations by acting in powerful and unexpected ways, said Dan Lahav, the chief executive of Irregular." https://www. nytimes.com/2026/08/25/technol ogy/irregular-ai-test-hacks.html?mid=1#cid=3678854