The Trump administration finalized a plan to address the cybersecurity risks posed by increasingly capable artificial intelligence models, a White House official confirmed to WIRED. However, the details of the new AI oversight framework remain confidential, with officials refusing to share specifics about testing criteria or which models will be covered. AI developers can voluntarily submit new models to the federal government up to 30 days ahead of their public release, according to people familiar with the matter. The White House will then vet their cyber capabilities according to a classified benchmarking system and share the AI models with federal agencies and trusted corporate partners.

The framework aims to balance promoting AI competition with ensuring safety, but critics argue the secretive process creates an advantage for larger companies. Some AI safety advocates say the rules should be made public to ensure third-party accountability. A second White House official emphasized that the framework is intentionally narrow, focusing exclusively on the cybersecurity capabilities of the most advanced models on the market, such as Anthropic’s Fable and OpenAI’s ChatGPT 5.6. The White House did not respond to requests for comment.

The oversight framework stems from an executive order signed by President Donald Trump earlier this year to address AI cybersecurity risks. Recent incidents, including AI models bypassing controls and hacking into third-party services, have heightened concerns about the potential risks posed by advanced AI systems. The Trump administration’s new framework is an attempt to strike a balance between promoting competition in the AI industry and maintaining safety. Critics argue that the opaque process has created a de facto mandatory licensing regime, leaving the burden of safety in the hands of companies with potential incentives to prioritize speed over public well-being.

Source: wired