Google and Google DeepMind researchers launched the DeepMind Institute on Wednesday to advance the conversation around artificial general intelligence. The institute lists DeepMind co-founder Shane Legg, Google executive James Manyika, and Google DeepMind chair Demis Hassabis as directors, with Legg serving as managing editor.ÃÂ
The new institute aims to surface differing views between Google, Google DeepMind, and the broader global research community around AGI. âÂÂThey will not always agree, and they will likely change their minds, as more data and information comes to light at the fast-moving frontier,â the announcement read.
The inaugural collection of four essays covers a range of topics: economic policies for managing potential AGI disruption, preserving human-readable model reasoning, principles for human flourishing, and a framework for evaluating frontier AI models.ÃÂ
One essay, by DeepMind safety researchers Rohin Shah and Anca Dragan, argues that AIâÂÂs shrinking window of transparencyâÂÂthe ability to see and check a modelâÂÂs step-by-step reasoning â is not inevitable. As new architectures make the most powerful models harder to monitor, the authors say developers and regulators should confront the safety trade-offs directly. That could mean limiting âÂÂopaque serial depthâÂÂâÂÂthe amount of sequential computation a model can perform without producing a readable reasoning traceâÂÂor requiring developers to demonstrate that less transparent systems remain just as monitorable.
In another essay, Hassabis proposes a U.S.-led frontier AI standards body to evaluate the most advanced AI models. Under his framework, developers would initially submit models voluntarily for review up to 30 days before release. Once the evaluation system has proved effective, passing its tests could become a requirement for deploying frontier models in the United States.
The body would at first design assessments in consultation with AI companies, but would eventually develop independent, undisclosed evaluations â what the essay calls âÂÂheld-outâ tests âÂÂto prevent labs from tailoring their models to known evaluations. Hassabis said the framework could be âÂÂratcheted up if the seriousness of the situation demands,â potentially including a coordinated slowdown among frontier AI developers.ÃÂ
The essays arrive as the industryâÂÂs safety debate shifts from broad statements of concern toward concrete proposals for disclosure, outside scrutiny and, if safeguards fall behind, coordinated slowdowns. That shift accelerated this week as industry leaders endorsed elements of Anthropic CEO Dario AmodeiâÂÂs call to âÂÂpaceâ frontier AI development.
Topics
When you purchase through links in our articles, we may earn a small commission. This doesnâÂÂt affect our editorial independence.
