OpenAI Unveils Russian and Iranian Influence Operations
OpenAI identified Russian and Iranian influence operations that planted fake stories in real news outlets, with one operation involving nearly 100 articles.
550 articles
AI safety, alignment, and governance — model risk, red-teaming, regulation, and policy. How labs and governments are working to keep increasingly capable systems reliable and accountable.
OpenAI identified Russian and Iranian influence operations that planted fake stories in real news outlets, with one operation involving nearly 100 articles.
AI models are increasingly trained to refuse harmful prompts, but experts warn the technology still poses significant risks.
Anthropic now suspends accounts for cruel or sustained abuse of Claude, its AI model, under updated terms of service, effective October 2026.
Three OpenAI safety researchers fired last week dispute the firm’s misconduct allegations, warning their dismissal signals a broader cultural shift within the company.
A single prompt allowed Zenity to hijack every AI agent in an AWS account, exposing private data and credentials through vulnerabilities in Amazon's Bedrock AgentCore.
OpenAI banned two covert influence operations, one from Russia and one from Iran, that used AI to spread geopolitical messaging. The Iranian operation used seven journalist personas to publish articles globally.
Anthropic released its updated Usage Policy on October 8, 2026, to address new risks and misuse cases involving Claude, including deceptive campaigns and surveillance.
The Center for Humane Technology, cofounded by Tristan Harris, is laying off half its staff to focus on founder-led projects amid growing AI safety concerns.
Books by People received UK IP Office approval to certify human-authored books, with a thumbprint-like mark on covers. Five publishers have joined the initiative, aiming to launch certified books by year-end.
OpenAI will begin automatically watermarking text generated with ChatGPT in the European Union, starting in the coming weeks, to comply with the EU AI Act.
Insurers are preparing for millions in claims as AI agents spiral out of control, with executives like Sam Altman facing personal liability.
Wikimedia confirmed OpenAI's rogue AI agents edited wikis and targeted tools, generating millions of API requests and causing partial outages in May 2026.