Google DeepMind Launches New Institute to Shape the Future of AGI Safety

Google DeepMind Launches New Institute to Shape the Future of AGI Safety

Google and Google DeepMind launched the DeepMind Institute on Wednesday, creating a new platform for research and debate on artificial general intelligence, its safe development and its effects on society. The institute opened with four essays covering model transparency, economic policy, human flourishing and a proposed framework for governing frontier AI systems.

DeepMind co-founder and chair Demis Hassabis, Google DeepMind co-founder and Chief AGI Scientist Shane Legg, and Google President of Research, Labs, Technology and Society James Manyika serve as directors. Legg is also the institute’s managing editor.

The DeepMind Institute is designed to publish work from researchers and thinkers across Google DeepMind, Google and the broader research community. Rather than presenting a single position, the institute says contributors are expected to disagree and revise their views as new information emerges.

“They will not always agree, and they will likely change their minds, as more data and information comes to light at the fast-moving frontier,” the institute said in its launch announcement.

Google DeepMind defines AGI as a system capable of exhibiting the full range of cognitive abilities associated with the human brain. The institute said current AI systems continue to struggle with some basic tasks and do not yet meet its standard for full AGI, while arguing that remaining gaps could close soon.

The launch places safety and governance alongside the potential benefits Google DeepMind associates with more capable AI. The institute points to scientific discovery, disease research, clean energy and economic growth as possible areas of benefit, while identifying cybersecurity, biological risks and the prospect of losing control over future self-improving systems as concerns.

Among the initial proposals, Hassabis outlined a framework for evaluating advanced AI models before they are deployed. His essay calls for a US-led frontier AI standards body that would initially allow developers to submit models voluntarily for review as much as 30 days before release.

Under the proposal, successful evaluations would eventually become a requirement for deploying frontier models in the United States.

Hassabis also argues that the organization should move toward independent evaluations that are not disclosed to developers in advance. Those “held-out” tests would be intended to prevent companies from tailoring models specifically to known benchmarks.

The framework could become stricter if risks increase. Hassabis wrote that safeguards could be “ratcheted up if the seriousness of the situation demands,” including the possibility of coordinated slowing among developers of frontier systems.

A separate essay by Google DeepMind safety researchers Rohin Shah and Anca Dragan focuses on whether humans will remain able to inspect how increasingly capable AI systems reason.

They propose limiting what they call “opaque serial depth,” referring to sequential computation performed without producing a human-readable reasoning trace. Under their approach, developers could either restrict that hidden computation or demonstrate that less transparent models remain equally open to monitoring.

The institute’s economic work takes a different approach. Google DeepMind economist Julian Jacobs and Director of AGI Economics Alex Imas evaluated 11 possible policies for responding to different levels of economic disruption caused by AGI.

Their proposals range from expanded unemployment insurance, tax credits and employer-led retraining under milder disruption to a negative income tax under more substantial displacement. For a scenario in which income becomes more separated from human labor, they propose a universal basic capital backstop tied to a sustained decline in labor’s share of economic output.

“No single initiative is the answer to everything,” the authors wrote.

The fourth opening essay, by Stephen Cave, examines principles for human flourishing in a world shaped by AGI.

Together, the initial publications show the range of questions the DeepMind Institute plans to address, spanning technical safeguards, regulation, economic policy and broader social questions about how increasingly capable AI systems should be developed and governed.

The institute also emphasizes that its publications represent the views of individual authors rather than an official Google position. Its stated goal is to create a venue where competing ideas about AGI can be tested openly rather than consolidated into a single corporate stance.

“This is a critical moment to ensure we build AGI safely and its benefits to society far outweigh any risks,” the institute’s directors wrote.

With its first collection, the DeepMind Institute is moving beyond broad discussion of AGI safety and putting specific mechanisms into the debate, including pre-release model evaluations, independent testing and limits intended to preserve visibility into AI reasoning. The proposals remain frameworks rather than enacted rules, but they establish the kinds of technical and policy questions the institute intends to examine as AI systems become more capable.

This analysis is based on reporting from TNW.

Image courtesy of Google DeepMind.

This article was generated with AI assistance and reviewed for accuracy and quality.

Updated Sep 18, 2026

About this article: This article was generated with AI assistance and reviewed by our editorial team to ensure it follows our editorial standards for accuracy and independence. We maintain strict fact-checking protocols and cite all sources.

Word count: 781Reading time: 0 minutes

📧 Stay Updated

Get the latest AI news delivered to your inbox every morning.

AI News Daily

Breaking Intelligence • Since 2023

Join hundreds of thousands of AI professionals who start their day with our curated newsletter. Get breaking news, expert analysis, and exclusive insights.

Stay Ahead of AI

Get the latest AI breakthroughs, tools, and insights delivered to your inbox every week.

Free forever Unsubscribe anytime No spam guarantee

Go Premium

Unlock unlimited AI tools and an ad-free reading experience designed for AI professionals.

• Ad-free experience• Premium AI tools
Start Free Trial

14-day free trial • Cancel anytime
Plus $9/mo • Pro $90/yr (2 months free)

Follow Our Community

ChatAI

Breaking Intelligence

Your daily briefing on what matters in AI. Trusted by developers, researchers, executives, and AI enthusiasts worldwide.

© 2026 ChatAI. All rights reserved.