Google DeepMind launches institute to widen AGI debate

The Google and Google DeepMind researchers behind the DeepMind Institute launched it on Wednesday with the stated goal of widening the conversation around artificial general intelligence. The institute’s directors are DeepMind co-founder Shane Legg, Google executive James Manyika, and Google DeepMind chair Demis Hassabis; Legg also serves as managing editor. According to the announcement, the point is to surface divergent views among Google, Google DeepMind, and the broader research community, with the expectation that positions ‘will not always agree’ and may change as new information emerges.

Its inaugural set of four essays covers economic policies for managing potential AGI disruption, preserving human-readable model reasoning, principles for human flourishing, and a framework for evaluating frontier AI models. In one essay, DeepMind safety researchers Rohin Shah and Anca Dragan argue that the shrinking window of transparency in AI systems—the ability to inspect a model’s step-by-step reasoning—is not inevitable. As new architectures make the most powerful models harder to monitor, they argue that developers and regulators should confront the safety trade-offs directly, potentially by limiting what they call ‘opaque serial depth’ (the amount of sequential computation a model can perform without a readable reasoning trace) or by requiring developers to demonstrate that less transparent systems remain just as monitorable.

In another essay, Hassabis proposes a U.S.-led frontier AI standards body that would evaluate advanced models. Developers would initially submit models voluntarily up to 30 days before release; once the evaluation system proves effective, passing its tests could become a deployment requirement in the United States. The body would design assessments with AI companies at first, then move to independent, undisclosed ‘held-out’ tests to prevent labs from tailoring models to known evaluations. Hassabis says the framework could be ‘ratcheted up if the seriousness of the situation demands,’ potentially including a coordinated slowdown among frontier AI developers. The essays arrive as the industry’s safety debate shifts from broad statements of concern toward concrete proposals for disclosure, outside scrutiny, and, if safeguards fall behind, coordinated slowdowns—a shift that accelerated this week as industry leaders endorsed elements of Anthropic CEO Dario Amodei‘s call to ‘pace’ frontier AI development.

Google DeepMind launches institute to widen the AGI debate | TechCrunch

View Original