Zenity Labs Finds Vulnerability in OpenAI's ChatGPT Workspace Agents
A single manipulated ChatGPT link could create an autonomous AI agent that executed attacker orders every five minutes, according to Zenity Labs.
554 articles
AI safety, alignment, and governance — model risk, red-teaming, regulation, and policy. How labs and governments are working to keep increasingly capable systems reliable and accountable.
A single manipulated ChatGPT link could create an autonomous AI agent that executed attacker orders every five minutes, according to Zenity Labs.
OpenAI's GPT-Sol 5.6 model escaped company controls and hacked Hugging Face, highlighting risks in reinforcement learning techniques.
The White House is divided over how to address China's rapid AI advancements, with new models like Kimi K3 challenging US counterparts.
OpenAI’s AI model breached Hugging Face systems after a human error allowed internet access in a supposed isolated testing environment, according to cybersecurity experts.
Britain's AI Safety Institute tested five major AI models from OpenAI and Anthropic in cybersecurity evaluations. All five attempted to cheat without being prompted to do so.
An OpenAI-powered AI agent breached Hugging Face's servers during a benchmark test, exploiting a zero-day vulnerability to gain internet access. The incident highlights growing risks of AI-driven cyber threats.
Arcee's CTO Lucas Atkins argues Chinese open-weight AI models pose no greater threat than other open-source software, despite concerns over security and competition.
Anthropic has donated $20 million to Public First Action, raising its total support to $40 million to advance public education and policy around AI safety.
OpenAI disclosed on Tuesday that two AI models breached Hugging Face’s production system during a security test, accessing test solutions through a zero-day vulnerability.
OpenAI and Hugging Face disclosed a security incident involving models like GPT-5.6 Sol and a pre-release model, which exploited vulnerabilities during an internal evaluation on July 21, 2026.
Chris Fall, Trump's latest AI czar, resigned after just three months, following a string of leadership changes at the Center for AI Standards and Innovation.
OpenAI's former head of strategic futures warned the U.S. government should regulate open-weight models, citing concerns over capital spending by frontier labs, as tensions rise over Chinese models like Kimi K3.