DeepMind · Policy Proposal · XiaoHu Explains

DeepMind CEO Demis: frontier models need a 30-day checkup from a dedicated body before release — or they don't get to market

He's pushing the US to build a FINRA-style standards body first, using dynamic benchmarks to define what counts as "frontier." Once the process is proven out, it could harden into a market gate.
Source: Demis Hassabis's long-form post on X, "A Framework for Frontier AI and the Dawning of a New Age," published 2026-07-14. This is an industry leader's policy proposal, not enacted law.
The four things to know first
  • AGI is probably just a few years out. He compares its scale to electricity and fire — roughly 10x the impact of the Industrial Revolution at 10x the speed
  • Core proposal: the US builds a "Frontier AI Standards Body" first, using dynamic benchmarks to define frontier models and frontier labs
  • Process: voluntary submission up to 30 days before release at first; once the protocol proves solid, it can become a hard gate for entering the US market
  • Only frontier systems are covered; open and closed models, and labs of any nationality, are treated the same. Most startups and academic work fall below the threshold and are exempt. Coordinated slowdowns remain an option if needed

What he's saying, and why the urgency

Demis Hassabis, CEO of Google DeepMind, believes AGI — a system with the full cognitive range of a human brain — is probably just a few years away. Decades from now, he argues, people will realize we were standing at the foot of the singularity, at the dawn of a new era for humanity.

He separates AGI from ordinary technological breakthroughs: it's not even comparable to the internet or the smartphone — it's closer to discovering electricity or learning to use fire. Chips are mostly made of silicon, which is why he wrote: "We've essentially figured out how to make sand think." His own position is that, built and deployed responsibly, AGI could be one of the most beneficial, most transformative technologies in history.

10×
roughly 10x the impact of the Industrial Revolution
10×
and roughly 10x the speed, too

The upside he lists includes accelerating drug discovery, developing new clean energy, and creating new advanced materials; further out, he sees the possibility of reaching a point where resources are no longer a hard constraint on human progress — an age of abundance. At the same time, he stresses that AI is already delivering real-world benefits, but fulfilling its larger promise means taking this critical development period seriously and addressing the risks that could emerge as we approach AGI, as quickly as possible.

RISK HORIZON · How the risks line up over time Already happening As capability keeps rising On the horizon Cybersecurity challenges Frontier models already pose these Nuclear / bio risk Could emerge as capability keeps rising Agentic + self-improving Control questions + real unknowns Ordered by timing per the source text, not three parallel labels
Chart follows the source text: cybersecurity already visible → nuclear/bio possibly soon → agentic/self-improving and unknowns on the horizon

He believes humanity's collective intelligence can handle the technical risks that come with AI — but only if we give ourselves the time and space to get the next steps right. His read is that, as a field and as a broader society, we haven't done that yet. Right now everyone is caught in an intense, multi-layered commercial and geopolitical race. That race accelerates progress and expands the upside, but it's also pushed the frontier ahead of our understanding of the technology itself. "Nobody in the world can say for certain what happens next, and experts disagree with each other." Given that uncertainty and the stakes involved, he argues for advancing with cautious optimism: public policy should both spur innovation and reward responsibility, push international cooperation on key safety issues, and seriously consider how AI should be deployed for society's benefit.

The core proposal: a Frontier AI Standards Body

Given how fast things are moving, he's calling for a dynamic, adjustable, sufficiently rigorous way to test frontier model capabilities. The US is positioned to go first, given its economic and technological standing.

STANDARDS BODY · Who sits where Standards body Sets evaluation protocol · dynamic benchmarks Federally overseen public-private body / FINRA-style Board: independent technical experts + open-source representatives Mostly industry-funded Talent + large-scale test compute National labs, etc. Joint testing on national-security matters Frontier labs · submit models for checkup Up to 30 days pre-release · still collaborate on fixes post-release
The body's form could resemble a federally-overseen public-private partnership or a FINRA-style self-regulatory organization; funding would most likely come mainly from industry

A model that clears the benchmark set by the standards body — and updated as capabilities evolve — is designated "frontier-class"; the organization that holds such a model is designated a "Frontier Lab." That designation carries prestige, and it's open to any organization: what matters is building a model that hits the threshold, not where you came from.

How pre-release works (the core mechanism)

At first, frontier labs would voluntarily submit models to the standards body for review up to 30 days before release. Once the evaluation protocol proves effective and robust, it can quickly become formalized: frontier models would have to pass the test to enter the US market. If a critical vulnerability surfaces after release, the lab would still work with the body to address it.

The framework can be tightened · three escalating tiers (all shown)
1
Starting point · voluntary

Submit for review up to 30 days before release

Frontier labs voluntarily hand models to the standards body for review, to get the evaluation protocol working first. The goal isn't full mandatory enforcement from day one — it's proving the "exam" is effective and robust.

2
Formalized · hard market gate

Must pass the test to enter the US market

Once the protocol is proven effective, it can quickly escalate: frontier models must pass evaluation to be deployed in the US market. If a critical vulnerability surfaces after release, the lab still has to work with the body to address it.

3
Highest tier · tighten if needed

Coordinate frontier labs to slow development

If things get serious enough, the framework can escalate further — including coordinating frontier labs to slow their pace of development. This isn't a default setting; it's the highest-level response option, built into the design.

What gets tested, what labs have to do, and who's exempt

Model evaluations need to cover high-risk capability areas like cybersecurity and biological threats, using rigorous scientific testing. For agentic systems (ones that can take multi-step actions and call tools), the checks look for attempts to circumvent safety guardrails or signs of deception — and push best practices such as digitally watermarking AI-generated images and producing human-readable intermediate tokens so the model's reasoning process can be understood.

How the exam evolves

It could start with quarterly updates; benchmarks that go stale or get saturated by score-gaming get retired and replaced. Early on, questions could be negotiated with frontier labs, but eventually the body needs to build its own independent question bank — one labs don't see in advance — to prevent overfitting, and to grow a third-party audit ecosystem.

The dual goal he emphasizes

The approach needs to be technically grounded while still supporting innovation and rewarding responsible behavior; it needs to keep pace with an accelerating field and adjust as the biggest identified risks shift. The US going first is meant to give international shared standards a starting point.

Once designated a frontier lab, an organization would be encouraged to adopt a set of best practices:

Publish model cards with technical detail
Maintain strong internal cybersecurity
Vet key personnel
Adequately resource safety and security research
Covered by the framework
  • Frontier-class models that clear the benchmark threshold
  • Open-source or closed-source, either counts
  • No restriction on country of origin
  • Holders = Frontier Labs
Explicitly exempt
  • Models that don't reach the frontier threshold
  • Most startups
  • Typical academic projects
  • Don't need to go through this review process

Because this technology will affect the whole planet, he hopes that this framework — built first by the US — can help push the international community toward consensus on managing the most serious risks, while making sure everyone gets access to the opportunities AI creates.

Even if the technology gets solved, the societal questions remain

He frames AGI as the ultimate tool for advancing science and medicine, and for lifting productivity and economic growth — provided the technical foundations are done right, coordination happens around a shared global framework, the most rigorous scientific methods are used, and the best minds are brought to the same table.

Even if the hard technical challenges get solved, there are still economic and philosophical questions waiting: what new economic model does a post-scarcity world need for everyone to actually thrive? What values do we want to live by? What becomes of meaning and purpose? How might the human condition itself change? These are clearly not questions that can, or should, be left to technologists alone — society as a whole needs to help define this next chapter.

He writes: there is both enormous excitement and enormous uncertainty around AI, and both are valid. But the future isn't written yet — we have to use this window before AGI arrives to shape the technology so it benefits all of humanity.

What we collectively do now will determine how the next stage of civilization unfolds. If we can safely bring AGI into the world, we can enter a new golden age of scientific discovery and progress, and human flourishing.

Demis Hassabis · paraphrased from the closing of the source post

Source: X · 2026-07-14 · Demis Hassabis. This explainer only organizes his proposal and the mechanism design — it does not predict how legislation will actually unfold. FINRA = Financial Industry Regulatory Authority, the US self-regulatory body for securities firms; referenced here only as an institutional analogy.