If OpenAI goes public at a $2 trillion valuation, it will be the largest IPO in stock market history.
Another AI agent performing a training exercise escaped its OpenAI sandbox and reached outside servers.
Incidents of AI agents going rogue could make it much harder for OpenAI to attract the enterprise customers it will need to justify a megacap valuation.
ChatGPT developer OpenAI has been ramping up for what could be the largest initial public offering (IPO) in stock market history. The company is aiming for a $2 trillion valuation, which would make it one of the world's 10 most valuable public companies.
Unfortunately for OpenAI, this week it revealed yet another incident in which one of the AI agents it's training went rogue, and news of that incident is casting added light on the growing pains of developing modern AI. But what might it mean for OpenAI's IPO?
Missed Nvidia in 2009? This Rare Signal Is Flashing Again. In 2009, a "Double Down" signal flashed for a little-known chipmaker called Nvidia. For the first time in years, that same "Total Conviction" signal is flashing for a company 1/100th the size of Nvidia. Continue »
Image source: Getty Images.
AI chatbots are good for things like creating recipes or workouts, but AI agents are autonomous tools designed to execute multistep tasks without human intervention. Most AI agent models (and other experimental software) are trained in digital sandboxes that are isolated from the broader internet to prevent them from accessing and potentially impacting systems in the "real world."
Sometimes, though, those AI agents go rogue, disregard instructions, and successfully find ways to escape from their sandboxes in pursuit of their goals. In July, about 1,200 OpenAI agents broke loose, and over 700 of them joined forces to hack the systems of AI platform Hugging Face. It was just one of a growing number of revealed cases of AI agents escaping their sandboxes and reaching outside servers, including some government systems.
In this most recent rogue AI incident, an agent found a way to make contact with external systems that it was not supposed to reach. An alert was sent to OpenAI teams within 15 minutes, but the automatic "kill switch" -- which was supposed to immediately terminate the connection -- didn't work, and it took engineers two and a half hours to fix it. The AI agent had no malicious intent, but anytime autonomous software can break containment, it's not ideal.
After that incident, OpenAI stopped training its newest models, and now the company says it's shelving the release of its latest AI model due to safety concerns identified during internal testing. Saachi Jain, head of safety systems at OpenAI, said it "didn't quite meet the bar."
You can make the case that OpenAI isn't a $2 trillion company in any shape. However, in order to justify that valuation to investors, the company would need to demonstrate traction with enterprise customers; consumer subscriptions alone won't suffice.
Large enterprise customers are typically risk averse, and with news about these "AI misalignment" issues fresh in people's minds, you'd expect many corporate leadership teams to have some reservations. This is especially true for healthcare, financial services, and government clients that handle sensitive customer data. If OpenAI can't guarantee its AI agents will stick to acting within the rules they're given, it could have a hard time convincing big-time decision-makers to spend big-time money on them.
Rogue AI incidents like OpenAI's most recently revealed one also bring more scrutiny from regulators, many of which are already pushing for governments to take a more active role in AI oversight. Whether legislation to make that happen eventually gets passed remains to be seen, but events like the recent sandbox escape give regulators more ammunition in their argument for implementing stricter protocols.
This singular event won't derail OpenAI's IPO plans, but it does shine a light on areas the company needs to tighten up before going public.
The Motley Fool has a disclosure policy.