NEW YORK: US artificial intelligence giant OpenAI promised on Wednesday to more systemically report instances of its models going off track, while also publishing six new reports on previously undisclosed incidents of AI misbehavior.
The transparency pledge follows a series of incidents at the company that have gradually come to light since July. The most serious involved two OpenAI models that, during testing, spontaneously broke out of their contained environment to access the internet and break into several websites and platforms.
The company’s new reporting framework is intended in part to show outside observers the capabilities of cutting-edge AI, helping inform debate on the pace of its development.
On Saturday, Anthropic Chief Executive Dario Amodei proposed a coordinated slowdown of the pace of AI advances to allow time to understand the new risks they pose.
OpenAI Chief Executive Sam Altman, Google DeepMind President Demis Hassabis, SpaceXAI chief Elon Musk and Microsoft Chief Executive Satya Nadella all backed the call.
In its announcement, OpenAI said, “We do not believe that the AI industry has solved alignment and monitoring to a sufficient degree to continue responsibly scaling at maximum speed for much longer.” “Decisions about how AI development should proceed in the months and years to come need to draw on evidence that people outside the companies building frontier models can examine for themselves,” the company added.
The company will now report problems involving, among other things, unauthorized actions by AI, escapes from oversight and spontaneous coordination between AI systems.
Published in Dawn, September 18th, 2026
































