OpenAIOpenAI NewsJun 23, 2026, 1:00 PM

Helping build shared standards for advanced AI

A condensed section focused on the key takeaways first.

Original Post

Quick Digest

Summary

A condensed section focused on the key takeaways first.

openaienmodel: gpt-5-mini-2025-08-07

Helping build shared standards for advanced AI

Key Points

  • Appia to define open, modular AI assessment specs
  • Evaluations must include structured disclosures and validation checks
  • Goal: interoperable, reusable evidence across institutions

Summary

OpenAI announced support for the Appia Foundation (hosted by the Linux Foundation) to create open, modular specifications that translate international standards and established frameworks into practical assessment criteria for AI across the value chain. Appia aims to establish a reusable trust layer so third parties can check conformity, produce comparable evidence, and enable national and international institutions to rely on each other’s technical findings. This effort complements OpenAI’s internal Preparedness and Frontier Governance frameworks and its engagement in ISO/IEC, NIST, and other standards forums.

Key Points

  • Appia will produce open, modular specifications to operationalize standards into concrete assessment criteria for models, infrastructure, and applications.
  • Third-party evaluations should disclose: the system tested, tool access and evaluation harness, elicitation methods, resources available, and validation checks performed.
  • The specification work is intended to make assessment practices interoperable across organizations, jurisdictions, and supply chains so governments and assessors can share trusted evidence and coordinate responses.
  • This builds on OpenAI’s Preparedness Framework and Frontier Governance Framework, which describe internal controls and regulatory-focused obligations (risk assessment, reporting, security controls, incident response, external review).
  • Engineers should design systems for auditability and structured evidence to align with emerging Appia specs and related standards activities.

Actionable guidance for engineers

  • Instrument models, tool access, and evaluation harnesses to capture reproducible logs and metadata.
  • Record and publish structured evaluation evidence: what was tested, how it was elicited, what resources were available, and what validation checks were run.
  • Track and contribute to Appia, ISO/IEC JTC1/SC42, NIST-led efforts, and related forums to adopt and help shape interoperable assessment practices.

Full Translation

Translations

A translation section that keeps the flow of the original article.

openaijamodel: gpt-5-mini-2025-08-07

先進的AIのための共有基準の構築を支援する

概要

能力が向上したモデルは、サイバー防御の強化、科学的発見の加速、専門知識へのアクセス拡大など多くの恩恵をもたらします。一方で、能力が誤解されたり、安全対策が不十分だったり、政府が適切に対応するための情報を欠いていたりすると、安全性やセキュリティ上のリスクも生じます。これらの恩恵を安全かつ自信を持って実現するには、ますます高機能化するシステムを評価・保護・統治するための技術的およびガバナンス上の能力を備えた制度が必要です。

Appia Foundation と共有技術言語の構築

このために、OpenAI は Linux Foundation がホストする Appia Foundation(別ウィンドウで開く)設立に協力しました。Appia は、国際規格や既存のフレームワークを AI バリューチェーン全体で実用的な評価基準に翻訳することを意図した、オープンでモジュラーな仕様を策定します。これにより、第三者が規格への適合性を検証し、モデル、インフラ、アプリケーションが異なる組織によって開発された場合でも、より明確で再利用可能な証拠を生成するという、重要な信頼レイヤーを構築することが期待されます。

Appia の取り組みは、各国・国際機関が互いの作業を信頼できるようにする共有の技術言語を生み出す助けになります。これは、高度な AI システムに対応するために必要な制度、基準、評価慣行を強化するための広範な取り組みの重要な次の一歩だと考えています。

ガバナンスと国際協力の必要性

我々の最近の「blueprint for democratic governance of frontier AI(前線的 AI の民主的ガバナンスに関する設計図)」は、その作業のロードマップを示しています。そこで提案している主な項目は次の通りです。

  • 耐久性のある米国の枠組み
  • 強化された Center for AI Standards and Innovation(CAISI)
  • 政府全体にわたるより広いレジリエンス戦略

この設計図は、前線的(frontier)リスクが国際的な範囲を有することも認めています。各国は互換性のある安全フレームワーク、リスク発見を共有するための信頼できるチャネル、そしてインシデントへの協調的対応をともに構築すべきです。国内の能力と国際協力は互いに補完し合う必要があります。

強力な機関(例:CAISI)は技術的専門知識を育成し、前線的システムを評価し、独立した評価エコシステムを支援できます。そうした有能な国レベルの機関のネットワークは、共有された手法を確立し、信頼される証拠を認識し、政府が共同行動を取るために必要な共通の技術理解を提供します。

評価慣行と OpenAI の取り組み

規格はこの努力の中心であり、信頼できる評価実践と技術的厳密さに基づく必要があります。我々がまとめた「trustworthy third-party evaluations(信頼できる第三者評価)の共通手引き」では、前線的評価がますます開示すべき項目として以下を挙げています。

  • テストされたシステム
  • ツールアクセスと評価ハーネス
  • 能力を引き出すために用いた手法
  • 利用可能なリソース
  • 結果を検証するために実施したチェック

また、米国の CAISI や英国の AISI とのテストパートナーシップを通じて、これらの原則を実践に移しています。これらの機関による前線能力評価や生物悪用対策に関する作業は、我々のシステムに具体的な改善をもたらしました。こうした取り組みは、性能を比較可能な形で検査できるよう標準化可能な慣行の基盤を作るという重要な役割を果たします。

これらの慣行は OpenAI のより広い安全インフラと補完関係にあります。主な内部・公開ドキュメントとしては次のものがあります。

  • Preparedness Framework(最も深刻なリスクを管理するための定義と運用化の基盤)
  • Frontier Governance Framework(リスク評価、モデル報告、セキュリティ管理、インシデント対応、外部専門家の意見導入などの具体的な規制上の義務に焦点を当てた公開ガバナンス文書)

これらの成果物は、広範な公約を検証・改善可能な運用的慣行に翻訳するのに役立ちます。Appia の作業は次の課題、すなわちそれらの慣行を組織、管轄、サプライチェーンを越えて相互運用可能にすることを目的としています。

OpenAI の参加と協働するフォーラム

OpenAI は既に、標準化とプレ標準化(pre-standardization)のより広いエコシステムに貢献しています。参加・協力している主なフォーラムは次のとおりです。

  • International Organization for Standardization and International Electrotechnical Commission Joint Technical Committee 1, Subcommittee 42 on Artificial Intelligence(ISO/IEC JTC 1/SC 42)(別ウィンドウで開く)
  • National Institute of Standards and Technology-led Artificial Intelligence Consortium(NIST 主導の Artificial Intelligence Consortium)(別ウィンドウで開く)
  • Frontier Model Forum(創設に協力)
  • Linux Foundation’s Agentic Artificial Intelligence Foundation(創設に協力)(別ウィンドウで開く)
  • Coalition for Secure Artificial Intelligence(参加)
  • Steering committee of the Coalition for Content Provenance and Authenticity(運営委員会に参加)(別ウィンドウで開く)
  • Internet Engineering Task Force(IETF)(別ウィンドウで開く)
  • Fast Identity Online Alliance(FIDO Alliance)(別ウィンドウで開く)

これらのフォーラムを通じ、そして今は Appia を通じて、我々の目的は前線開発から得られた教訓をオープンで技術的に裏付けられた実践へと翻訳し、各国や企業、独立評価者が管轄を超えて利用できるようにすることです。


著者: OpenAI 公開日: 2026-06-23

Helping build shared standards for advanced AI | OpenAI News | DocsDigest