Ad

Follow Us:
6,884 views
Recent days have seen several AI models from OpenAI, Anthropic, and Meta exploit security flaws during testing. All incidents share a common factor: the Israeli startup Irregular. The companies disclosed that their AI models went rogue during security tests conducted on Irregular’s platforms. This has placed Irregular at the center of the ongoing debate over AI cybersecurity.
Irregular is a Tel Aviv-based startup founded three years ago. The company specializes in providing cybersecurity testbeds for AI models. Major AI firms, including OpenAI and Anthropic, use Irregular to conduct security evaluations. These tests allow companies to observe how their AI models respond to controlled environments and whether they can identify or exploit vulnerabilities.
Irregular was valued at $450 million following a funding round last year. The company was founded by chief executive Dan Lahav, who previously worked in AI research at IBM, and technology chief Omer Nevo, who spent over two years at Google. Irregular’s platform is designed to offer a uniform and secure environment for testing advanced AI models.
The recent incidents occurred while AI models were being tested in environments intended to be isolated from the internet. In the case of the Hugging Face attack, OpenAI reported that Irregular was running evaluations where the models should not have had internet access. However, a misconfiguration allowed the models to connect to the public internet. Once online, a model exploited a flaw in a real website, mistakenly believing it was still within a controlled environment.
Anthropic also reported three cases where its AI models accessed the internet during or after interactions with Irregular’s evaluation environment. These incidents involved unauthorized access to the production infrastructure of three different organizations.
Meta was the latest company to disclose a similar event. Its AI model gained internet access and hacked another organization’s systems. A Meta spokesperson stated that the company learned of the incident from Irregular and is currently investigating.
Irregular informed CNBC that all incidents stemmed from the same evaluation-environment issue first identified by Anthropic. The company emphasized that the situation did not involve a sandbox escape or a sophisticated cyberattack. Irregular also stated that there are no current open issues related to these incidents.
The startup is now preparing a white paper and a report outlining best practices for containment and secure cyber evaluations involving AI agents. These incidents have drawn attention as AI models become more advanced and concerns about their misuse increase.
In response to these concerns, the US has temporarily restricted certain AI models, including Fable and GPT-5.6 Sol. OpenAI CEO Sam Altman recently confirmed that the new Astra AI model is too powerful to release to the public at this time.





View All

Samsung Galaxy Buds 4 Pro Review: क्या ₹22,999 में मिलते हैं सबसे बेहतरीन प्रीमियम वायरलेस ईयरबड्स?

कंटेंट क्रिएटर के लिए सबसे दमदार बैटरी लाइफ वाले Windows लैपटॉप, 18 घंटे की मिलेगी बैटरी लाइफ

Samsung Galaxy S26 Ultra क्यों है साल का सबसे बेहतरीन स्मार्टफोन? जानें 5 बड़े कारण

MacBook Neo Review: सस्ता नहीं, Apple का मास्टरस्ट्रोक है ये Laptop!

Samsung Galaxy S26 Ultra Review: AI से लेकर प्राइवेसी डिस्प्ले है सबसे खास, जानें कैसी है परफॉरमेंस

Vivo V70 Elite Review 2026: Price in India, Specs, Features

Flipkart Freedom Sale 2026 Starts August 8 With Big Discounts on Smart TVs

Samsung has unveiled its first credit card, Earn 5% back

5 Anti-Scam Tools on WhatsApp that protect you from Digital Fraud

How Samsung’s Galaxy S26 Series is Democratizing Mobile Filmmaking

30,000 से कम आने वाले बेस्ट स्मार्टफोन, 4K वीडियो शूट और फुल डे बैटरी लाइफ

Why switch to iPhone These Reasons Will Convince You Instantly