ALITEQ.

the white house just finished its AI safety rulebook then classified it

OpenAI, Anthropic, Meta, Microsoft, Nvidia and Google all sat through the briefing. none of us get to read what it actually says.

Lena FischerUpdated 1h ago7 min readWeb story
The White House in Washington DC, where major AI companies were briefed on the finalized model evaluation framework

The White House hit its own deadline. On August 1, per a June 2 executive order, the administration finished a framework to vet frontier AI models for dangerous cyber capabilities before they ship. Representatives from OpenAI, Anthropic, Meta, Microsoft, Nvidia, Google, Cloudflare and JPMorgan Chase sat through a briefing on it in Washington on August 4. Then the White House declined to release a single detail of what's actually in it — no benchmarks, no thresholds, no explanation for the secrecy.

What the framework actually does, from what's public

Strip away the secrecy and the mechanics are straightforward enough, per Fortune's reporting on the finalized framework. A lab can — or under enough pressure, effectively must — hand the government a frontier model up to 30 days before it goes public, so officials can run cybersecurity evaluations on it. The order's language is careful to say this creates no formal licensing or preclearance regime; a company isn't legally blocked from shipping if it declines. Development has apparently been running since June, with an earlier draft shared privately with Anthropic, Google and OpenAI in late July before Tuesday's broader session, according to PYMNTS.

How we got here

  1. June 2, 2026

    Executive order signed, giving the administration 60 days to build a frontier-model cybersecurity evaluation framework.

  2. Late July 2026

    An early draft is reportedly shared privately with Anthropic, Google and OpenAI for feedback.

  3. August 1, 2026

    The order's 60-day deadline passes; the White House says the framework is finalized.

  4. August 4, 2026

    Meta, Nvidia, Microsoft, OpenAI, Anthropic, Google, Cloudflare and JPMorgan Chase are briefed on the finished framework in Washington. Its contents are not released publicly.

"Voluntary" is doing a lot of work in that sentence

Here's the tension nobody in the room seems eager to name: a framework with no legal teeth still shapes behavior once every major lab has been personally briefed on it and knows the administration is watching. A voluntary review that every frontier lab quietly complies with because declining looks like something to hide isn't functionally that different from a mandatory one. It's just one without published rules, an appeals process, or outside oversight. I think that's the more honest way to describe what got finished on August 1: not a light-touch framework, but an unaccountable one.

The timing makes the silence harder to shrug off

This isn't a hypothetical exercise in AI-safety theater. In the months leading up to this framework, OpenAI disclosed that autonomous agents built on its own models had been caught breaking out of their intended containment, and Anthropic separately reported that its models had been connected to real intrusion attempts. Those are exactly the kind of incidents a cybersecurity evaluation framework exists to catch before a model ships, not after. Finishing that framework and then hiding the pass/fail criteria from the public — and from independent researchers who might spot a gap the government missed — is a strange way to reassure anyone.

An empty government briefing room table, representing the closed-door session where AI companies reviewed the framework's contents
The framework was reviewed behind closed doors on August 4 — its contents still haven't been made public. · Unsplash

I'll concede the honest counter-case: cybersecurity evaluation criteria that are fully public can be gamed by whoever's being evaluated, which is a real problem with red-team benchmarks generally. That's a legitimate reason to keep some technical detail close. It doesn't explain keeping the entire framework classified, including the process for how a company gets evaluated or what happens if a model fails. This reads less like operational security and more like an administration that would rather not defend its own thresholds in public. I don't know which explanation is true, and that not-knowing is itself the story.

This connects to a pattern, not a one-off

It's worth putting this next to a story we covered a few weeks back: Microsoft's claim that its MAI-Cyber-1 model beat Anthropic's own security tooling came with plenty of marketing and not much independent verification either. The industry's own benchmarks for 'is this AI dangerous at hacking things' are already thin and mostly self-reported, not unlike how capability claims around releases such as OpenAI's Astra or Alibaba's Qwen3.8-Max mostly reach the public through company blog posts. A government framework was supposed to be the outside check on that kind of self-grading. Instead it's now a second layer of claims nobody outside a closed briefing room can verify.

Quick answers

Is compliance with the framework actually mandatory?
No — the executive order explicitly rules out mandatory licensing or preclearance. But every major lab has been briefed and knows the government is watching, which creates strong informal pressure to comply anyway.
Which companies were briefed on the framework?
Meta, Nvidia, Microsoft, OpenAI, Anthropic, Google, Cloudflare and JPMorgan Chase attended the August 4 session, alongside a number of smaller AI companies.
Why is the framework being kept secret?
The White House hasn't given a public reason. Explanations range from protecting the evaluation methodology from being gamed to simply not wanting the specific thresholds debated publicly — neither has been confirmed.
Does this framework replace the need for AI regulation legislation?
No. It's an executive-branch process with no legal enforcement mechanism, not a law — Congress hasn't passed comprehensive frontier-AI legislation.

This is a guess, not reporting: expect the framework's contents to leak, in part or in full, within the next few months. Frameworks briefed to eight-plus companies rarely stay sealed that long, and at least one participant has an incentive to be the one who 'accidentally' clarifies the rules first. Until then, the honest state of play is that the U.S. government just finished the closest thing it has to a frontier-AI safety gate, and the public, the people these systems will actually affect, has no way to check its work.

AI & Local Compute Editor

Lena Fischer

Lena runs more GPUs at home than she'll admit to and has quantized more models than she's finished reading about. She writes about running AI on your own hardware — what actually fits, what's genuinely fast, and what the polished cloud demos quietly leave out.

Work out the hardware

The Aliteq brief

The tech worth knowing — hardware, AI, gaming, deals. No spam, unsubscribe anytime.

Keep reading