Where Elegance Meets Intelligence

THE AI STREET JOURNAL

← Back to the paper

Policy 17 September 2 min read

DeepMind institute opens with proposal for US frontier model tests

Demis Hassabis proposes outside assessment before advanced models are released. His framework would start voluntarily, with mandatory tests a possible later step.

An editor checks a proof beside an old printing press in the morning light.
Editorial illustration · The AI Street Journal

The DeepMind Institute begins with four essays covering economic policy, readable model reasoning, human flourishing and frontier-model evaluation. Aditya Mehta reported the launch for TechCrunch.

Its directors are DeepMind co-founder Shane Legg, Google executive James Manyika and Google DeepMind chair Demis Hassabis. Legg also serves as managing editor. The institute aims to publish differing views, rather than a single agreed position.

Voluntary first, potentially compulsory later

Hassabis proposes a US-led standards body to assess the most advanced models. Developers would initially submit them voluntarily, up to 30 days before release. Once the system had proved effective, passing its tests could become a condition of deployment in the United States.

Assessments would initially be designed with AI companies. The body would later develop independent, undisclosed tests to stop developers tailoring models to known examinations. Hassabis also leaves room for a coordinated slowdown if the risks warrant it.

Keeping reasoning inspectable

In a separate essay, safety researchers Rohin Shah and Anca Dragan propose confronting the trade-offs of less transparent models. Options include limiting sequential computation without readable reasoning, or requiring evidence that opaque systems remain equally monitorable. These are proposals, not adopted requirements.

If adopted, the testing framework would add an external assessment before frontier-model deployment. Its move towards undisclosed tests addresses a specific weakness: a model can perform well on familiar evaluations without demonstrating equally reliable behaviour elsewhere.

Sources & publication notes

Published in our 18/09/2026 edition. Source dates are shown above.

The AI Street Journal · Free to read

Make this your morning paper.

Three AI stories, clearly explained and illustrated. Free in your inbox.

Prefer to listen? Choose your podcast preferences →