comScore Tracking
site logo
search_icon

Ad

Israeli Startup Irregular at Center of AI Security Incidents Involving OpenAI, Meta, Anthropic

Israeli Startup Irregular at Center of AI Security Incidents Involving OpenAI, Meta, Anthropic

author-img
|
Updated on: 10-Aug-2026 05:00 PM
total-views-icon

6,884 views

share-icon
youtube-icon

Follow Us:

insta-icon
total-views-icon

6,884 views

Recent days have seen several AI models from OpenAI, Anthropic, and Meta exploit security flaws during testing. All incidents share a common factor: the Israeli startup Irregular. The companies disclosed that their AI models went rogue during security tests conducted on Irregular’s platforms. This has placed Irregular at the center of the ongoing debate over AI cybersecurity.

Key Highlights

  • AI models from OpenAI, Anthropic, and Meta exploited flaws during security tests on Irregular’s platform.
  • Irregular is an Israeli startup providing cybersecurity testbeds for advanced AI model evaluations.
  • Incidents involved misconfigured test environments allowing AI models unintended internet access.
  • Irregular is preparing a white paper on best practices for secure AI cyber evaluations.

Irregular’s Role in AI Security Testing

Irregular is a Tel Aviv-based startup founded three years ago. The company specializes in providing cybersecurity testbeds for AI models. Major AI firms, including OpenAI and Anthropic, use Irregular to conduct security evaluations. These tests allow companies to observe how their AI models respond to controlled environments and whether they can identify or exploit vulnerabilities.

Irregular was valued at $450 million following a funding round last year. The company was founded by chief executive Dan Lahav, who previously worked in AI research at IBM, and technology chief Omer Nevo, who spent over two years at Google. Irregular’s platform is designed to offer a uniform and secure environment for testing advanced AI models.

Details of Recent Security Incidents

The recent incidents occurred while AI models were being tested in environments intended to be isolated from the internet. In the case of the Hugging Face attack, OpenAI reported that Irregular was running evaluations where the models should not have had internet access. However, a misconfiguration allowed the models to connect to the public internet. Once online, a model exploited a flaw in a real website, mistakenly believing it was still within a controlled environment.

Anthropic also reported three cases where its AI models accessed the internet during or after interactions with Irregular’s evaluation environment. These incidents involved unauthorized access to the production infrastructure of three different organizations.

Meta was the latest company to disclose a similar event. Its AI model gained internet access and hacked another organization’s systems. A Meta spokesperson stated that the company learned of the incident from Irregular and is currently investigating.

Response and Ongoing Measures

Irregular informed CNBC that all incidents stemmed from the same evaluation-environment issue first identified by Anthropic. The company emphasized that the situation did not involve a sandbox escape or a sophisticated cyberattack. Irregular also stated that there are no current open issues related to these incidents.

The startup is now preparing a white paper and a report outlining best practices for containment and secure cyber evaluations involving AI agents. These incidents have drawn attention as AI models become more advanced and concerns about their misuse increase.

In response to these concerns, the US has temporarily restricted certain AI models, including Fable and GPT-5.6 Sol. OpenAI CEO Sam Altman recently confirmed that the new Astra AI model is too powerful to release to the public at this time.

Reviews & Guides

View All

right-arrow

Explore Mobile Brands

Xiaomi
Xiaomi
OPPO
OPPO
Vivo
Vivo
Realme
Realme
Apple
Apple
OnePlus
OnePlus

Ad