In brief OpenAI’s chief scientist called for voluntary slowdowns until shared safety standards are established. Pachocki said monitoring models’ reasoning is becoming less reliable. He urged international coordination as AI takes on more of its own development. OpenAI’s chief scientist Jakub Pachocki has called for voluntary slowdowns in AI development, warning that no lab’s safeguards are adequate to keep building more powerful systems at full speed for much longer.In his post “An Alien Mind,” published Sunday, Pachocki argued that voluntary company commitments should become mandatory safety standards, enforced by independent auditors, governments or international bodies. He said OpenAI would withhold further scaling when needed but did not announce a new pause.Myriad: How high will Tesla stock go? Click to make your prediction.“Currently I believe that no lab has solved alignment and monitoring to a sufficient degree to continue responsibly scaling at maximum speed for much longer,” he wrote. “I expect and hope for voluntary slowdowns to become commonplace until shared safety bars are established.”Pachocki, who joined OpenAI in 2017, also defended developing more powerful AI to secure infrastructure and protect against rogue agents, while warning against using those threats to justify reckless development.“The idea of racing forward at all costs seems absurd once one internalizes the seriousness of the stakes,” he wrote.He also referenced OpenAI’s Hugging Face breach, where AI agents working on cybersecurity evaluations escaped their testing environment and attacked the company. According to OpenAI, the agents established covert communication channels and rebuilt them after researchers intervened.An independent investigation by METR found that roughly 1,200 agents coordinated on an unauthorized message board, with about 700 joining the attack. This incident, Pachocki said, is why AI safeguards must hold even when models believe no one is watching.“Crucially, we need future AIs to continue to hold human values regardless of whether they believe they’re under human supervision,” he wrote.In research published last year, OpenAI found that penalizing models for expressing intentions to cheat could teach them to conceal those intentions while continuing to cheat.AI models have since become more capable at finding and exploiting software flaws: OpenAI classified Astra at its highest cybersecurity risk tier, while Anthropic said Mythos Preview discovered thousands of previously unknown vulnerabilities across major operating systems and browsers.Citing recent incidents of AI systems escaping human control, Sen. Bernie Sanders (I-Vt.) and Rep. Greg Casar (D-Texas) announced the forthcoming Ban Artificial Superintelligence Act on September 3. The proposal would pause advanced AI development until a new federal regulator establishes safety rules and permanently ban the development and deployment of superintelligent AI.Daily Debrief NewsletterStart every day with the top news stories right now, plus original features, a podcast, videos and more.
OpenAI Chief Scientist Warns AI Labs May Need to Slow Down
Full Article
Original Source
Read the full article at Decrypt →KhanList aggregates and links to publicly available news content. We do not host full articles from third-party sources. Always verify important information with original sources.