2026-08-04-Tue · Anthropic

From Issue 4 (2026-08-04) · 10 stories in this issue

❯ White House Says Advanced AI Model Evaluation Framework Completed on Schedule, But Details Withheld

ON TIME, NOT PUBLICA White House official said the voluntary evaluation framework for advanced AI models required by the June 2 executive order has been completed on schedule — but the White House has not disclosed the framework’s contents, who has reviewed it, or when companies will begin using it. According to Axios, a working-level meeting with companies is scheduled for Tuesday to introduce the finished product to vendors.

SECRECY CLAUSEThis opacity is not an ad hoc decision; it was written into the original text itself. The executive order explicitly states that the benchmarking process for assessing models’ advanced cyberattack capabilities is classified, and that the threshold determining which models fall under oversight is likewise classified. The White House says the industry partners it has engaged go well beyond OpenAI, Anthropic, and Google. Many had expected a discussable evaluation standard to be made public, since a “voluntary framework” typically derives its binding force from transparency — once companies sign on, outsiders can hold them to their commitments. What has emerged instead: a framework whose contents no one outside can see, paired with a pledge of voluntary participation.

THE DILEMMAThis design creates a dilemma for policy researchers and corporate compliance teams alike: the former cannot determine where the threshold is set or which models are covered, making it impossible to judge whether the framework is strict or lenient; the latter must decide how many resources to dedicate to aligning with a standard without seeing the full text. The most directly affected area is how model capabilities are publicly disclosed — if cyber capability evaluation results are themselves classified, the boundary of what vendors can and cannot write in release notes shifts accordingly. After Tuesday’s meeting, most likely only attendees will know what the framework looks like.

▪ SIGNALAn invisible framework — its binding force ultimately rests on whether signatories are willing to voluntarily disclose what they’ve done.