OpenAI Calls for Global Safety Standards to Govern Self-Improving AI Systems

OpenAI has released a framework urging the United States and global partners to build technical standards for frontier artificial intelligence. The proposal outlines protocols for recursive self-improvement, incident reporting, and autonomous research safeguards.

OpenAI global AI safety standards concept showing an international policy forum, AI globe, and safeguards for governing self-improving artificial intelligence systems, including ri
OpenAI has outlined a policy position advocating for US-led global standards to govern recursive self-improvement and automated research.

A Framework for Frontier Governance

OpenAI has published a policy framework titled "Building standards for the next phase of AI," advocating for international technical baselines to govern advanced model development. The proposal urges the United States to lead a global effort alongside international bodies and safety institutes. Rather than creating restrictive licensing regimes or pre-release approval bottlenecks, the initiative focuses on creating shared definitions for technical evidence and safety baselines.

The push addresses three major global challenges in artificial intelligence governance:

  • Regulatory Fragmentation: Incompatible definitions, metrics, and reporting requirements between countries make comparing technical evidence difficult.
  • Collective Action Pressures: Independent national efforts risk an uncoordinated acceleration that could outpace democratic oversight and safety evaluation.
  • Uneven Capability Distribution: Global disparities in technical expertise complicate joint risk assessments and international oversight.

Regulating Recursive Self-Improvement

Central to the document is the management of recursive self-improvement, a process where artificial intelligence models assist in creating successive generations of systems. OpenAI warns that fully autonomous self-improvement must not proceed without verified safety measures, as unmonitored development could cause humans to lose practical oversight over automated research processes.

To manage these capabilities safely, OpenAI proposes specific technical guidelines:

  • Human Review Thresholds: Defining specific operational triggers that require immediate human intervention during automated research.
  • Incident Reporting Frameworks: Standardising severity classifications and disclosures for alignment and security issues across laboratories.
  • International Collaboration: Utilizing entities such as AI safety institutes, ISO, and the Frontier Model Forum to establish non-discriminatory standards that apply equally to open-source and proprietary developers.

Get the next one by email