The White House has finalized a voluntary framework for reviewing advanced AI models before release, but the testing criteria are staying private. That turns frontier model launches into a new kind of prerelease security review, with open-weight models and outside researchers still sitting largely outside the process.
OpenAI paused some internal Astra work after early evaluations could not rule out “Critical” cyber capabilities. The move shows frontier model launches are becoming cybersecurity release-gate events, with stricter test environments, monitoring, government review, and third-party controls now part of the path to deployment.