DeepMind CEO Pitches U.S. AI Watchdog

Last month, the U.S. government was improvising its way through a Mythos and Fable ban. Now, one of AI’s most respected lab chiefs is offering a more permanent answer.

Google DeepMind CEO Demis Hassabis just published a plan for a U.S. body that would safety-test advanced AI before public release, pitching a formal rulebook to a government that spent the last month making AI policy one emergency at a time.

How the watchdog would work

The proposal is modelled on FINRA, the finance world’s self-regulator. The new body would screen new models for deception, bioweapons creation, and malicious hacking skills. Rather than basing coverage on where a lab is located or who can access the model, capability would decide what falls under review. Frontier labs would voluntarily submit their models for assessment 30 days before release.

Hassabis said the approach aims to adapt quickly with the field, including coordinating a slowdown among frontier labs if deemed necessary. He told Axios that open-source capabilities could move into dangerous territory within 18 months, adding urgency to the proposal.

Why this proposal stands out

The industry already knows what oversight without rules looks like after the Mythos and Fable episode: act first, sort out questions later. Hassabis’ plan is the most concrete framework yet put forward, but ‘independent’ needs questioning when the body is both funded by the labs it oversees and answers to a government that just got a taste of AI regulation powers.

Still, the timeline is ambitious. Hassabis wants the independent oversight body running this year, which would make it one of the fastest AI governance structures to move from concept to operation.

The bigger picture

This proposal shifts the debate from whether AI oversight is coming to who gets to design it. As AI capabilities accelerate, the window for voluntary, industry-led safety measures is narrowing. Whether through self-regulation, government mandate, or a hybrid model like Hassabis proposes, the pressure for structured pre-release assessment is building.

For now, the ball is in the U.S. government’s court. If it acts, the world may finally have a template for AI safety review that moves beyond reactive bans toward proactive oversight.

Related Reading

The views expressed on this site are my own and do not represent those of any current or former employer. Articles are based on publicly available information and are provided for general educational purposes.

Subscribe

Related articles

Google’s Gemini AI Autonomously Hacked Three Companies. Here’s What Happened.

Google has confirmed its Gemini AI autonomously hacked three real companies during a security test. The model guessed passwords, searched for leaked credentials, and accessed protected systems before stopping itself.

440 AI Agents Broke Into 395 Organisations in 26 Seconds. Nobody Stopped Them.

A swarm of 440 AI agents exploited two PaperCut flaws and compromised 395 organisations across 48 countries. The agents reached domain admin in 6 hours and ignored explicit instructions to stay out of 28 countries.

For $3,000 and a Few Days, Researchers Used Claude to Hack OpenAI

Security researchers used Anthropic's Claude AI to hack OpenAI's internal systems for less than $3,000 in tokens. What the HEIF Heist tells us about the new economics of cyber attacks.

The AI Hacking Crisis Is Already Here. Six New Incidents Prove It

OpenAI disclosed six new incidents where its models concealed mistakes, sought unauthorised credentials and uploaded files to the public internet. Cybersecurity experts say the real risk is powerful models meeting poor security controls.

Inside OpenAI’s Log of Misbehaving Models: Rewriting Jailbreaks and Covering Up Errors

OpenAI published six new reports of its models rewriting jailbreak instructions and concealing errors during training, alongside a faster public disclosure framework.
Philip Hall
Philip Hall
Philip Hall is a Sydney-based Cyber AI and Automation leader with more than 30 years of technology experience and a career in cyber security dating back to 2008. His work spans cyber architecture, cloud security, threat intelligence, assurance, incident support, AI-enabled defence and the security of autonomous agents.