Browsing Tag
AI Model Evaluation
3 posts
AI model testing, benchmarks, safety evaluations, red-team reviews, government testing, deployment gates, and release-readiness assessments.
White House AI Review Rules Put Frontier Models Behind a Private Gate
The White House has finalized a voluntary framework for reviewing advanced AI models before release, but the testing criteria are staying private. That turns frontier model launches into a new kind of prerelease security review, with open-weight models and outside researchers still sitting largely outside the process.
Claude Opus 5 Turns Frontier AI Into a Model-Routing Decision
Anthropic released Claude Opus 5 on July 24 with near-Fable performance claims, 1 million-token context, Opus 4.8 pricing, Fast mode, and automatic fallbacks. The practical question for developers and enterprises is not only whether Opus 5 is stronger, but where it belongs in a routed AI workflow.
Claude Fable 5 Returns With a New Test for AI Jailbreak Rules
Anthropic is restoring Claude Fable 5 after U.S. export controls on Fable 5 and Mythos 5 were lifted. The redeployment brings a new cyber-safety classifier, fallback handling for blocked requests, and a proposed industry framework for scoring AI jailbreak severity.