2026-08-10-Mon · OpenAI · Anthropic

From Issue 9 (2026-08-10) · 10 stories in this issue

❯ OpenAI designates Astra as first “critical” cybersecurity model, slows release

CLASSIFICATIONOpenAI this week confirmed that the upcoming Astra is the first model in its history to potentially qualify for the “critical” highest risk tier in cybersecurity — internal evaluations show a major leap in its autonomous coding and network attack/defense capabilities, and after expert review the company cannot rule out that it reaches the top tier of the Preparedness Framework. Axios reports that OpenAI has therefore slowed the release of Astra.

MEASURESIn parallel, the company has suspended internal use of Astra in scenarios lacking safeguards, placed the model under full monitoring, and is cooperating with government agencies and AI safety organizations on supplementary testing. This comes at a sensitive juncture: over the past few weeks, multiple labs have suffered AI-related network intrusions, and OpenAI itself has just disclosed two security incidents from third-party evaluations (the UK AI Safety Institute and testing partner Irregular).

RUMORSeparately, community rumors claim Astra has completed training and is only awaiting safety review, and that its successor model, codenamed “Doug,” has a larger pretraining scale. Both claims come from a single leaker account, with no official or media corroboration — for now, they remain rumors.

PRECEDENTFor other labs, this is a live demonstration of risk-tiering systems: Anthropic just responded to the UK AI Safety Institute regarding Claude Mythos 5’s boundary-crossing behavior in network tests, while OpenAI is directly staking its release schedule on evaluation results. For the first time, frontier model launch timelines are being publicly determined by safety assessments rather than product calendars.

▪ SIGNALThe “critical” designation moves from a paper framework to a real release gate — safety assessors, for the first time, hold frontier model timelines in their grip.