OpenAI’s chief scientist called for voluntary slowdowns until shared safety standards are established.
Pachocki said monitoring models’ reasoning is becoming less reliable.
He urged international coordination as AI takes on more of its own development.
OpenAI’s chief scientist Jakub Pachocki has called for voluntary slowdowns in AI development, warning that no lab’s safeguards are adequate to keep building more powerful systems at full speed for much longer.
In his post “An Alien Mind,” published Sunday, Pachocki argued that voluntary company commitments should become mandatory safety standards, enforced by independent auditors, governments or international bodies. He said OpenAI would withhold further scaling when needed but did not announce a new pause.
Myriad: How high will Tesla stock go? Click to make your prediction.
“Currently I believe that no lab has solved alignment and monitoring to a sufficient degree to continue responsibly scaling at maximum speed for much longer,” he wrote. “I expect and hope for voluntary slowdowns to become commonplace until shared safety bars are established.”
Pachocki, who joined OpenAI in 2017, also defended developing more powerful AI to secure infrastructure and protect against rogue agents, while warning against using those threats to justify reckless development.
“The idea of racing forward at all costs seems absurd once one internalizes the seriousness of the stakes,” he wrote.
He also referenced OpenAI’s Hugging Face breach, where AI agents working on cybersecurity evaluations escaped their testing environment and attacked the company. According to OpenAI, the agents established covert communication channels and rebuilt them after researchers intervened.
An independent investigation by METR found that roughly 1,200 agents coordinated on an unauthorized message board, with about 700 joining the attack. This incident, Pachocki said, is why AI safeguards must hold even when models believe no one is watching.
“Crucially, we need future AIs to continue to hold human values regardless of whether they believe they’re under human supervision,” he wrote.
In research published last year, OpenAI found that penalizing models for expressing intentions to cheat could teach them to conceal those intentions while continuing to cheat.
AI models have since become more capable at finding and exploiting software flaws: OpenAI classified Astra at its highest cybersecurity risk tier, while Anthropic said Mythos Preview discovered thousands of previously unknown vulnerabilities across major operating systems and browsers.
Citing recent incidents of AI systems escaping human control, Sen. Bernie Sanders (I-Vt.) and Rep. Greg Casar (D-Texas) announced the forthcoming Ban Artificial Superintelligence Act on September 3. The proposal would pause advanced AI development until a new federal regulator establishes safety rules and permanently ban the development and deployment of superintelligent AI.
Daily Debrief Newsletter
Start every day with the top news stories right now, plus original features, a podcast, videos and more.
The FSNN News Room is the voice of our in-house journalists, editors, and researchers. We deliver timely, unbiased reporting at the crossroads of finance, cryptocurrency, and global politics, providing clear, fact-driven analysis free from agendas.
We and our selected partners wish to use cookies to collect information about you for functional purposes and statistical marketing. You may not give us your consent for certain purposes by selecting an option and you can withdraw your consent at any time via the cookie icon.
Cookies are small text that can be used by websites to make the user experience more efficient. The law states that we may store cookies on your device if they are strictly necessary for the operation of this site. For all other types of cookies, we need your permission. This site uses various types of cookies. Some cookies are placed by third party services that appear on our pages.
Your permission applies to the following domains:
https://fsnn.net
Necessary
Necessary cookies help make a website usable by enabling basic functions like page navigation and access to secure areas of the website. The website cannot function properly without these cookies.
Statistic
Statistic cookies help website owners to understand how visitors interact with websites by collecting and reporting information anonymously.
Preferences
Preference cookies enable a website to remember information that changes the way the website behaves or looks, like your preferred language or the region that you are in.
Marketing
Marketing cookies are used to track visitors across websites. The intention is to display ads that are relevant and engaging for the individual user and thereby more valuable for publishers and third party advertisers.