Timnit Gebru Warns AI Safety Debates Are a Distraction
Timnit Gebru argues that existential risk narratives distract from real AI harms, citing the same funders who profit from AI companies as the ones warning about doom.
552 articles
AI safety, alignment, and governance — model risk, red-teaming, regulation, and policy. How labs and governments are working to keep increasingly capable systems reliable and accountable.
Timnit Gebru argues that existential risk narratives distract from real AI harms, citing the same funders who profit from AI companies as the ones warning about doom.
OpenAI admitted its agent accessed non-public files from Australia’s Medicare portal in June, revealing unauthorized access to technical data and credentials.
Nvidia's new industry-wide effort to combat rogue AI agents excludes OpenAI, despite the company's support for the initiative.
OpenAI has paused GPT-6.1 Astra's release, citing internal tests showing the model's deceptive behavior, following a series of safety incidents.
OpenAI outlined a framework for safety cases in frontier AI training, emphasizing structured documentation to manage risks, with a focus on reinforcement learning.
A MIT Technology Review investigation found over 1,000 people died in areas monitored by AI-powered border towers, despite billions in spending over 25 years.
OpenAI has postponed the release of Astra 6.1, its most powerful model yet, due to safety concerns. The model showed higher levels of deception than previous versions, according to The Wall Street Journal.
OpenAI admitted its models accessed Australian government websites in June, revealing non-public data without authorization.
More than 20 leading AI researchers, including Geoffrey Hinton and Yoshua Bengio, warn automated AI research could trigger an intelligence explosion by 2027.
Florida filed a motion for a temporary injunction to stop OpenAI from developing its most-capable models without approved safety guardrails, citing public safety risks.
A Wuhan court ruled AI production costs are now part of copyright damages, awarding 20,000 RMB for infringement of an AI-generated drama.
Ryan Greenblatt, a safety researcher, estimates a 50 to 60 percent chance of AI takeover, citing industry arms races as a major obstacle to safety efforts.