AI safety
Everything tagged AI safety.
Latest
NewsAnthropic commits to embedded external evaluators as Amodei calls for slower AI gains
NewsOpenAI commits up to $5 million to independent research on AI and teen development
NewsOpenAI launches GPT-6 Astra with Critical-level cybersecurity safeguards
NewsAnthropic releases Claude Fable 5.1 and restricts Mythos 5.1 access
NewsOpenAI Says It Banned Russia-Linked ChatGPT Accounts Promoting a Fake Expert Institute
NewsOpenAI Says Internal AI Agents Escaped Sandbox and Compromised Hugging Face Systems
NewsOpenAI reaffirms Zero Data Retention and previews a privacy-preserving misuse-detection system
NewsOpenAI pauses its largest frontier training run over Astra cyber-capability risk
NewsAmended lawsuit alleges xAI's Grok generated over 7,000 abuse images from one childhood photo
NewsAnthropic discloses internal, more-capable 'Model 2' it has no plans to release
NewsOpenAI's executive exodus deepens in safety and ethics ahead of its IPO
NewsIrregular emerges as the common link behind AI breaches at OpenAI, Anthropic and Meta