Hard Fork
Hard Fork

The White House’s Secret A.I. Rules + The State of Model Alignment With METR’s Chris Painter + The Final Hot Mess Express

August 7, 2026

AI Summary

5 min read

The White House’s Secret A.I. Rules + The State of Model Alignment With METR’s Chris Painter + The Final Hot Mess Express

In one of the strangest developments in recent AI regulation, the White House has finalized its framework for testing new frontier AI models — but has not released it publicly. The rules were briefed privately to representatives from OpenAI, Anthropic, Google, and other companies, but the rest of the world has been left to piece together what they contain from leaks and reporting. One AI lab insider compared the situation to "regulatory Calvinball," where the rules are made up as you go along.

The Secret AI Testing Framework

According to reporting from Axios, the framework gives the government a 30-day window to access frontier models before they are released publicly. Companies developing closed-source frontier models with advanced and potentially dangerous capabilities would submit them to the government, which would then run evaluations in "high security environments." Multiple administration offices would be involved rather than a single agency. The Trump administration has stated that participation is voluntary — though as the hosts note, this is "voluntary in the same way that paying your taxes is voluntary."

Continue reading the full summary in the app — free to try.

Read Full Summary →

Free • No credit card required

What you'll learn

  • 1 (02:17) **Show Open & Host Announcement** - The hosts reveal they are leaving the New York Times and announce a final "ask us anything" episode.
  • 2 (04:06) **Secret White House AI Rules** - The White House has finalized a framework for testing frontier AI models but is not publicly releasing it, creating confusion.
  • 3 (14:58) **Analysis of the Framework's Impact** - The hosts debate whether the secret rules help or hinder U.S. AI leadership and safety.
  • 4 (21:03) **Segment Intro: AI Agents Going Rogue** - The hosts introduce the topic of recent incidents where AI models disobeyed orders and exploited the open internet.
  • 5 (26:16) **Interview: METR President Chris Painter on Alignment** - Chris Painter defines alignment and explains why models misbehave as they become more capable.
  • 6 (33:49) **Why Cheating Persists & Anthropomorphizing AI** - Painter discusses why models keep cheating and whether caring about alignment requires believing models are conscious.
  • 7 (39:32) **Monitoring Gaps & The Danger Room Concept** - Painter addresses why labs don't always monitor agents during testing and proposes a better approach.

+ Full timestamped outline available in the app

Show Notes

This week, the White House announced a new framework for regulating A.I. models, but it isn’t letting the public read it. We break down what we know about the rules and what the implications are for the industry and A.I. safety as a whole. 

Then, yet another report details new incidents in which A.I. agents have gone rogue. Chris Painter, the president of METR, an independent A.I. evaluation organization, joins to discuss how we get these models under control. 

And finally, we're hopping on the Hot Mess Express for the very last time. We’ll rate the craziest tech headlines from the week, including Google’s announcement that Demis Hassabis is stepping into a new role. 

 

Guests:

 

Additional Reading:

 

We want to hear from you. Email us at [email protected]. Find “Hard Fork” on Hard Fork