Back to archive
AI·Article

Governments and OpenAI push independent frontier-AI assessment while common rules remain unsettled

Leaders and senior officials from 20 countries plus the European Commission have called for mandatory pre-deployment testing, independent evaluation and shared reporting of serious frontier-AI incidents. OpenAI has now separately committed to deep-access third-party assessments across training, evaluation and deployment, but neither track creates an adopted common standard or binding international verification regime.

Published 22 Sept 2026, 09:49 · Updated 23 Sept 2026, 15:37

Story history · 1 earlier version

22 Sept 2026, 09:49

OpenAI proposes frontier-AI standards while cross-lab safety talks continue

Governments turn the standards debate into a multilateral call

A group of leaders and senior officials from 20 countries, together with European Commission President Ursula von der Leyen, published a joint call on 21 September for tighter oversight of frontier AI. The statement asks companies to use transparent safety protocols with mandatory pre-deployment testing and independent evaluation, asks governments and regional organisations to coordinate common standards and shared reporting of serious safety incidents, and asks UN member states to explore an international institution able to set standards, enable verification and convene states when capability thresholds are crossed.

The signatories include leaders or senior officials from Norway, Finland, Australia, Bahrain, Canada, Denmark, Estonia, Germany, Iceland, Ireland, Kazakhstan, Kenya, Latvia, Moldova, the Netherlands, Singapore, Spain, South Africa, Türkiye and the United Arab Emirates. The text remains open for further endorsements. Its material effect today is political rather than legal: the declaration creates no binding testing requirement, regulator or approval process by itself.

OpenAI and Anthropic supply different technical baselines

OpenAI’s 21 September standards paper proposes shared measurements of frontier capability progress and research autonomy, human-review triggers and common incident classification and reporting. Reuters independently reported the proposal. OpenAI presents the United States as a preferred leader of that standards effort, but its paper does not establish adoption by governments, standards bodies or rival labs.

Anthropic’s 17 September R&D Automation Index measures a different part of the same governance problem. Anthropic says Claude led 26% of the AI research-and-development work it measured in August and collaborated on more than 90%, with no measured subset fully autonomous. Those are confirmed internal measurements rather than a cross-lab benchmark: Anthropic used its own Claude-based agents and judges, notes there is no common industry methodology, and says independent third-party evaluation is still being set up.

OpenAI sets out a deep-access assessment model

On 22 September, OpenAI said it is committed to supporting independent assessments with deep access across training, evaluation and deployment. It proposes four priority areas: safety cases, critical safeguards, capability and alignment evaluations, and independent investigation of critical misalignment incidents. The company says assessors should be able to challenge assumptions, identify missed risks and reach their own conclusions about safeguards. This is a first-party commitment and proposed assessment model, rather than evidence that every OpenAI system has already undergone such scrutiny.

OpenAI’s proposed principles call for clearly scoped and pre-registered claims, proportionate access within legal, security and intellectual-property constraints, transparent methods and uncertainty, assessor expertise and conflict disclosure, security and confidentiality protections, actionable findings and evidence-based publication. That direction overlaps with the governments’ demand for qualified independent evaluators with sufficient access. Anthropic and Accenture’s separate embedded-evaluator programme remains a company-funded arrangement, so the three efforts should not be treated as one verification system.

The operating rules are still unsettled

Reuters and TechCrunch reported on 15 September that OpenAI, Anthropic and Google DeepMind had been discussing AI safety across company lines for several weeks, but those talks have produced no public joint charter, common evaluation protocol, standing institution or binding deployment rule. Europe’s role is becoming more concrete through the new declaration and planned frontier-lab discussions, while French and German officials have rejected broad calls to halt AI progress. Nothing reviewed here shows the three labs jointly endorsing the government statement, OpenAI’s standards structure or its new third-party assessment principles.

What would change the evidence status

The evidence remains mixed across independently announced tracks. The government statement is an official multilateral declaration, and OpenAI’s assessment commitment and Anthropic’s measurements are established as first-party evidence, but there is still no adopted common frontier-AI standard. A formal UN or recognised standards process, binding national or EU rules, a published cross-lab protocol, or independent assessments demonstrating comparable methods across labs would materially strengthen the evidence for a unified governance framework.

Source trail

18 sources · 10 primary · 8 reference

01
European Commission Audiovisual Service — State of the European Union 2026
Primary · 16 Sept 2026, 09:00
https://audiovisual.ec.europa.eu/en/stories/M-010025
02
Reuters — EU to invite frontier labs for AI-risk talks
Reference · 16 Sept 2026, 09:47
https://www.reuters.com/world/eus-von-der-leyen-invite-frontier-labs-talks-tackling-ai-risks-2026-09-16/
03
Reuters — OpenAI working with Anthropic and Google on AI safety
Reference · 15 Sept 2026, 16:50
https://www.reuters.com/technology/openai-is-working-with-anthropic-google-ai-safety-bloomberg-news-reports-2026-09-15/
04
Dario Amodei — We Must Pace the Frontier
Primary · 12 Sept 2026, 02:00
https://darioamodei.com/post/we-must-pace-the-frontier
05
Reuters — French finance minister says slowdown calls favour US AI leaders
Reference · 16 Sept 2026, 17:21
https://www.reuters.com/business/finance/calls-slow-down-ai-development-serve-interests-us-ai-leaders-says-french-finance-2026-09-16/
06
Reuters — Germany says halting AI development is not viable
Reference · 14 Sept 2026, 13:32
https://www.reuters.com/legal/litigation/germany-says-halting-ai-development-not-viable-calls-us-china-involvement-2026-09-14/
07
Mark Zuckerberg (@finkd) — on independent lab pacing
Primary · 16 Sept 2026, 01:01
https://x.com/finkd/status/2099997096896274533
08
Reuters — Zuckerberg says AI labs have enough incentive to build safely
Reference · 16 Sept 2026, 04:21
https://www.reuters.com/business/metas-zuckerberg-says-ai-labs-have-enough-incentive-build-safely-2026-09-16/
09
Anthropic — Partnering with Accenture on embedded evaluation
Primary · 18 Sept 2026, 02:00
https://www.anthropic.com/news/accenture-embedded-evaluation
10
Accenture — Embedded evaluators at Anthropic
Primary · 18 Sept 2026, 02:00
https://newsroom.accenture.com/news/2026/accenture-and-anthropic-partner-to-build-team-of-embedded-evaluators-at-anthropic
11
TechCrunch
Reference · 15 Sept 2026, 17:47
https://techcrunch.com/2026/09/15/openai-anthropic-google-have-been-in-talks-on-ai-safety-for-weeks/
12
OpenAI - Building standards for the next phase of AI
Primary · 21 Sept 2026, 02:00
https://openai.com/index/building-standards-next-phase-ai/
13
Reuters - OpenAI calls for U.S.-led global frontier-AI technical standards
Reference · 21 Sept 2026, 19:02
https://www.reuters.com/legal/government/openai-calls-us-take-lead-global-efforts-develop-technical-standards-2026-09-21/
14
Anthropic Institute - Measuring the pace of AI development
Primary · 17 Sept 2026, 02:00
https://www.anthropic.com/institute/measuring-pace-of-ai-development
15
Government of the Netherlands — A Call for Control of Frontier AI Models
Primary · 22 Sept 2026, 02:00
https://www.government.nl/documents/2026/09/22/a-call-for-control-of-frontier-ai-models
16
President of the Republic of Finland — A Call for Control of Frontier AI Models
Primary · 21 Sept 2026, 02:00
https://www.presidentti.fi/en/a-call-for-control-of-frontier-ai-models/
17
Hacker News discussion
Reference · 22 Sept 2026, 02:00
https://news.ycombinator.com/item?id=49806145
18
OpenAI - Priorities and principles for effective third party assessments
Primary · 22 Sept 2026, 02:00
https://openai.com/index/priorities-principles-third-party-assessments/