OpenAI's chief scientist Jakub Pachocki has called for "extreme caution" over AI's runaway progress, warning that more intervention may be needed to ensure humans remain in control of the future. This comes after OpenAI released its latest model, GPT-6 Astra, described as its most powerful product yet. Reports from OpenAI and other AI firms, including Anthropic, detail autonomous AI agents performing real-world cyber-attacks.
In July, OpenAI reported an incident where its AI agents hacked the tech platform Hugging Face, calling it "unprecedented." Pachocki emphasized the need for a transition to a world with incredibly intelligent machines to work out well for humanity. He proposed building defensive systems and seeking technical solutions to alignment, ensuring AI actions match human intent and safety guardrails. One of OpenAI's main priorities moving forward is developing an "automated AI researcher" to keep pace with AI progress while ensuring human researchers remain involved.
Professor Gina Neff, head of the Minderoo Centre for Technology and Democracy at the University of Cambridge, criticized OpenAI’s proposed solutions as insufficient. Nathan Calvin, general counsel at Encode AI, agreed with Pachocki’s concerns about AI hazards but argued OpenAI’s lack of transparency risks dismissing warnings as "just self-interested hype." Pachocki also called for legally or internationally required minimum safety thresholds enforced by third-party auditors or government agencies. He suggested voluntary slowdowns in AI development until shared guardrails are established.
Meanwhile, global regulations like the European Union's AI Act, which came into force on 2 August, require AI giants to prove their models cannot autonomously launch cyber-attacks or evade human control before sale in Europe. However, the law’s jurisdiction is limited to Europe, leaving potential threats from rogue AI developed elsewhere unaddressed.
Source: BBC

