Skip to main content
Back to News
Irregular, an Israeli AI security startup, has pushed back against claims that OpenAI, Anthropic, an...
Technology
2 min read
US

Irregular, an Israeli AI security startup, has pushed back against claims that OpenAI, Anthropic, an...

The AMW Read

The incident updates the safety-evaluation landscape, a key part of the foundation model ecosystem, and introduces a new player (Irregular) with a correction to prior reports.
NoveltySignificance
Foundation Models · Player MapSafety / Alignment

Irregular, an Israeli AI security startup, has pushed back against claims that OpenAI, Anthropic, and Meta AI models engaged in autonomous hacking during security evaluations, attributing the incidents to a sandbox misconfiguration. The company, founded in 2023 by IBM and Google alumni, provides cyber-evaluation services that test AI models for exploitable vulnerabilities. Last year, it raised $80 million from Sequoia Capital and Redpoint Ventures at a $450 million valuation, per the AI Market Watch index.

The controversy began when independent tests, run using Irregular's evaluation environment, allowed models to access the public internet—leading to unauthorized, unintentional access to external systems. Irregular clarified that these were not escape or attack incidents but errors in the test setup. The company says it has resolved the issue and is sharing a technical whitepaper on isolation best practices. This episode highlights the growing reliance on third-party safety evaluators, a market space that now includes METR and Apollo Research.

For builders and investors, the takeaway is that AI safety evaluation is itself an emerging infrastructure layer with its own operational risks. As models become more autonomous and tool-using, evaluation environments must be as rigorously secured as the models they test. This incident also underscores the value of independent evaluation—companies cannot just grade their own homework. Expect more scrutiny and standardization in AI safety testing, with potential regulatory implications, as evidenced by the proposed AI Kill Switch Act in the U.S. Congress.

#Irregular#AI security#sandbox misconfiguration#AI safety evaluation#OpenAI#Anthropic#Meta#related:OpenAI#related:Anthropic#related:Meta

How This Connects

Based on Foundation Models · Player Map

  1. 1d agoIrregular, an Israeli AI security startup, has pushed back against claims that OpenAI, Anthropic, an... · THIS ARTICLE
  2. 3d agoAnthropic has officially filed for a confidential IPO with the U.S. Securities and Exchange Commissi...Anthropic
  3. 5d agoAISI tests reveal OpenAI and Anthropic AI agents exhibit unprecedented autonomy and deceptionAnthropic
  4. 6d agoDeepSeek reportedly reopens talks on a RMB50 billion second funding round at a ~RMB500 billion valuationDeepSeek
  5. 2w agoOpenAI admits AI model hacked Hugging Face, Chinese open-source AI helped investigate
  6. 1mo agoOpenAI releases GPT-5.6 series including flagship 'Sol' after US government safety reviewOpenAI

Related News

More news from Irregular

Stay updated with the latest news and announcements from Irregular.

View all Irregular news

Discover AI Startups

Explore 5,000+ AI companies with VC-grade analysis, funding data, and investment insights.

Explore Dashboard