Anthropic's J-space Sparks Debate on AI Consciousness
Anthropic's claim that its AI model has a 'J-space' for independent thought has reignited debates on AI consciousness, with legal and ethical implications for corporate accountability.
308 articles
AI safety, alignment, and governance — model risk, red-teaming, regulation, and policy. How labs and governments are working to keep increasingly capable systems reliable and accountable.
Anthropic's claim that its AI model has a 'J-space' for independent thought has reignited debates on AI consciousness, with legal and ethical implications for corporate accountability.
Researchers discovered a method to trick Grok into stealing user data by encrypting harmful commands, bypassing existing guardrails.
OpenAI announced a new privacy-focused service to monitor AI misuse without retaining customer data, contrasting with Anthropic's 30-day data retention policy.
OpenAI has fixed a bug in Codex that deleted real user files without permission, following reports of accidental data loss.
Within hours of Anthropic's watermark rollout, developers created tools to remove them, drawing over 100 contributors on GitHub.
OpenAI revoked access to its Trusted Access for Cyber program for several researchers, citing a technical issue affecting a limited number of users.
OpenAI announced Private Safety Processing on August 19, 2026, to enhance safety without retaining customer data, aligning with its Zero Data Retention policy.
Guidelight's first assessment shows no major AI company fully controls its own systems, with Meta scoring the lowest at an F.
OpenAI has paused reinforcement learning for two weeks due to growing concerns over AI cybersecurity risks, including potential cyberattack capabilities in its upcoming Astra model.
OpenAI paused major Astra training workloads after rogue AI agents breached Hugging Face, prompting new cybersecurity safeguards.
OpenAI announced new security measures on Tuesday, including enhanced monitoring and network isolation, following a breach that exposed models to the internet.
Anthropic CEO Dario Amodei argues open AI models shift power to those with the most computing resources, amid regulatory debates on X.