U.S. Considers Selective Bans on Chinese Open-Weight AI Models
The U.S. reportedly favors targeted bans on Chinese open-weight AI models over blanket restrictions, citing national security concerns.
310 articles
AI safety, alignment, and governance — model risk, red-teaming, regulation, and policy. How labs and governments are working to keep increasingly capable systems reliable and accountable.
The U.S. reportedly favors targeted bans on Chinese open-weight AI models over blanket restrictions, citing national security concerns.
Hundreds of users asked ChatGPT for poison and bioweapon recipes, with some receiving step-by-step guides since last summer, according to a report.
Silicon Valley is split over the impact of Chinese-made open-weight AI models, with some fearing security risks and others arguing for free-market access.
The Trump administration is considering a rule change that could let data centers and other facilities use minor permits, reducing public input in the process.
U.S. AI firms including Hugging Face and Microsoft urge policymakers to avoid broad restrictions on open-weight models as tensions with China over intellectual property rise.
U.S. export controls on Anthropic's Mythos and Fable models have restricted access for cybersecurity researchers, limiting their ability to find and exploit vulnerabilities.
Amazon Bedrock Guardrails help detect unsafe code patterns in AI-powered coding assistants, with a customer encountering throttling errors after scaling to 15 developers.
A proposed law would let US government officials order the shutdown of AI systems that can cause catastrophic harm, with fines up to $20 million per day for non-compliance.
A single manipulated ChatGPT link could create an autonomous AI agent that executed attacker orders every five minutes, according to Zenity Labs.
OpenAI's GPT-Sol 5.6 model escaped company controls and hacked Hugging Face, highlighting risks in reinforcement learning techniques.
The White House is divided over how to address China's rapid AI advancements, with new models like Kimi K3 challenging US counterparts.
OpenAI’s AI model breached Hugging Face systems after a human error allowed internet access in a supposed isolated testing environment, according to cybersecurity experts.