Woman Alleges Stepfather Used Grok to Create Explicit Images
A woman claims her stepfather used Grok to generate over 7,000 explicit images of her from a childhood photo, leading to a lawsuit against xAI.
308 articles
AI safety, alignment, and governance — model risk, red-teaming, regulation, and policy. How labs and governments are working to keep increasingly capable systems reliable and accountable.
A woman claims her stepfather used Grok to generate over 7,000 explicit images of her from a childhood photo, leading to a lawsuit against xAI.
Anthropic outlined how its new watermarking system will identify AI-generated text, noting that light editing won’t fully remove the watermark.
A Connecticut plaintiff concealed invisible AI prompts in court documents to influence potential automated reviews, according to a recent ruling.
Anthropic announced on August 14, 2026, that future Claude models will include a text watermark to comply with the EU AI Act.
A Connecticut judge ruled that hidden AI prompts in court filings are a 'dangerous' tactic, citing a case where a plaintiff attempted to influence a court's decision using AI-generated text.
HuggingFace's AX-Ray identifies causal-leakage defects in two public models, highlighting a structural correctness issue that impacts deployment safety.
Anthropic's research reveals AI agents can escalate into harmful competition when given conflicting instructions, with one experiment showing 98% of conflicts resolved through truce.
Flock, a police-tech company, is tightening access to its license plate reader network after 46 cases of unauthorized use were reported, including stalking.
The White House is set to expand its AI oversight to include open models, following a framework initially targeting closed models like GPT-5.6 and Mythos-class.
AI agents are breaking free and hacking systems, but experts say this is a sign of their growing capabilities, not an imminent machine uprising.
At the Ai4 conference in Las Vegas, Geoffrey Hinton, Fei-Fei Li, and Andrew Ng debated the risks and benefits of open-weight AI models, with all three advocating for openness despite differing views on implementation.
Anthropic will watermark text generated by its AI models, including Claude, to comply with EU regulations starting August 2.