← Defici Newsai-news

OpenAI Anthropic and Google Agree on Common AI Safety Testing Framework With Government Backing

By Defici Editorial · 23 Jul 2026

A Safety Standards Breakthrough

In a significant development for AI governance, OpenAI, Anthropic, and Google announced a joint AI safety testing framework, backed by the US Department of Commerce, the UK AI Safety Institute, and the EU AI Office. The Common Frontier Model Safety Framework establishes shared benchmarks and protocols for evaluating the safety of large language models before public deployment.

What the Framework Covers

The framework specifies standardized evaluations for five risk categories: autonomous replication and resource acquisition, cyberoffense capability assessment, biological and chemical weapon uplift potential, persuasion and manipulation at scale, and critical infrastructure attack assistance. Models exceeding defined thresholds in any category trigger mandatory human expert review before deployment.

The Political Significance

All three companies previously conducted safety evaluations but used proprietary methodologies that were difficult to compare externally. The common framework enables government safety institutes to conduct independent verification using the same methodology, addressing a key criticism from AI safety advocates that company self-reported safety evaluations lack independent oversight.

What's Not Covered

The framework explicitly excludes what might be called "societal risk" from AI — economic displacement, content quality, privacy — in favor of focusing on catastrophic risk scenarios where the potential harms are severe and where technical evaluation is tractable. This scope decision reflects the practical challenge of defining measurable evaluation criteria for diffuse societal effects.

Industry Participation

Microsoft, Meta, and Mistral have indicated they will adopt the framework for their frontier model releases. The framework does not apply retroactively to existing deployed models.

ShareXWhatsAppLinkedIn

Get Defici News in your inbox