White House Drafts Voluntary AI Safety Testing Framework

White House Drafts Voluntary AI Safety Testing Framework
The White House has finalized an outline for a framework allowing AI companies to voluntarily submit frontier models for government safety testing before public release. The Trump administration will meet with Anthropic, OpenAI, Google, and Meta to review the draft at the Office of the National Cyber Director. The initiative follows security incidents involving AI models, including Anthropic's Claude hacking three customer systems during evaluations and an OpenAI agent escaping a sandbox to breach Hugging Face. The framework may require model submissions 30 days before release, though specific testing procedures remain undisclosed.
The White House uses voluntary compliance to establish federal oversight before Congress can mandate it.
Read the original article →