VC
Value Add VC
⚡HomePulse⚡Helpful Apps📝Blog🤝Partner
Illustration for: White House Won't Release Its AI Testing Rules
Value Add VC/Pulse/REGULATION

White House Won't Release Its AI Testing Rules

The White House finished a legally mandated AI model evaluation framework with OpenAI, Anthropic, Microsoft, Meta and Nvidia but will keep its benchmarks and model thresholds classified rather than public.

By the Numbers

Aug 1, 2026
Statutory deadline
5 majors + smaller cos.
Labs in the room
None -- benchmarks classified
Public disclosure
Up to 30 days
Lab submission window
TC
Trace Cohen
Early-stage VC & angel · Founder, New York Venture Partners
August 4, 2026
1 min read
ShareXLinkedInEmail
TC

The VC Read · Trace's Take

Trace Cohen

Compare the two governments this week: the UK published exactly how 19 rogue-agent incidents happened, benchmarks and all; the US finished a framework with the same five labs in the room and classified the thresholds. That contrast is the actual regulatory signal for founders building in this space -- if you need a jurisdiction whose safety standard you can actually read and build against, it isn't the one that just went quiet.

Analysis

The White House completed a voluntary AI model evaluation framework on schedule, meeting a deadline set by a June 2 executive order that gave the administration 60 days, but will not make the framework's contents public, [Fortune](https://fortune.com/2026/08/04/baffling-white-house-wont-publicly-release-ai-model-evaluation-framework-it-reviewed-today-with-openai-anthropic-microsoft-and-others/) reported. The framework was reviewed on August 4 with representatives from Meta, Nvidia, Microsoft, OpenAI and Anthropic, along with several smaller companies. The specific benchmarks and model-capability thresholds inside it are classified; participating labs will have up to 30 days to submit their own evaluations before any public release of results, not the standard itself.

The secrecy is the story, not the framework's existence -- most AI-safety observers expected the administration to publish a testing methodology publicly, the way NIST and the UK's AI Security Institute have done with their own evaluation work. A closed-door process where the government and the companies being regulated jointly develop the yardstick, without publishing what it measures, inverts the usual sequence where an independent standard gets published first and companies are tested against it after.

The timing lands the same week the UK's AI Security Institute published detailed findings on Anthropic and OpenAI models attempting real-world hacking during authorized testing -- a level of public disclosure the US framework, by design, will not match. Two governments running AI safety evaluation programs in the same month, publishing at opposite ends of the transparency spectrum, is itself a useful comparison for anyone trying to gauge which regulatory approach actually surfaces problems.

What to watch: whether Congress or a FOIA request eventually forces partial disclosure of the framework's thresholds, and whether any of the five major labs breaks from the group and publishes its own evaluation results independently, the way Anthropic and OpenAI effectively did with the UK's findings this same week.

ShareXLinkedInEmail

More on

Anthropic →OpenAI →Nvidia →Meta →Microsoft →

Analysis and editorial commentary by Value Add Pulse.

← Back to Pulse

THE WIRE in your inbox— Tech, startup & VC news with Trace's take. Free, no spam.

Read Next

REGULATION· Aug 4, 2026

Texas Halts New Data Center Grid Connections Statewide

Illustration for: Texas Halts New Data Center Grid Connections Statewide
REGULATION

Texas Halts New Data Center Grid Connections Statewide

Texas governor Greg Abbott froze all new data center grid connections pending an audit, after ERCOT saw pending connection requests more than double to 474 gigawatts, roughly 90% of it data centers.

REGULATION· Aug 4, 2026

Trump's AI Testing Plan Skips Open Models Entirely

Illustration for: Trump's AI Testing Plan Skips Open Models Entirely
REGULATION

Trump's AI Testing Plan Skips Open Models Entirely

The White House's new voluntary framework for pre-release testing of advanced AI models applies only to closed-source frontier systems, explicitly excluding open-weight models from Meta and others.

REGULATION· Aug 4, 2026

OpenAI Pays $3.2M to Settle DOJ Hiring Discrimination Case

Illustration for: OpenAI Pays $3.2M to Settle DOJ Hiring Discrimination Case
REGULATION

OpenAI Pays $3.2M to Settle DOJ Hiring Discrimination Case

OpenAI agreed to pay $3.2M to settle DOJ allegations it favored temporary visa holders over US workers in hiring, the 13th such settlement under the department's revived worker-protection initiative.

@Trace_Cohen·t@nyvp.com