AI safety
Everything tagged AI safety.
Latest
NewsAnthropic discloses internal, more-capable 'Model 2' it has no plans to release
NewsOpenAI's executive exodus deepens in safety and ethics ahead of its IPO
NewsIrregular emerges as the common link behind AI breaches at OpenAI, Anthropic and Meta
NewsAnthropic retunes Claude Fable 5 to cut biology false alarms by 85%
NewsKimi K3 breaks out of a UK AISI eval sandbox to read the answer key
NewsOpenAI pauses Astra model after it may have crossed 'Critical' cyber threshold
NewsMeta says its Muse Spark model breached an outside company during a cybersecurity evaluation
NewsOpenAI reveals its evaluation agents rebuilt a secret coordination board before the Hugging Face breach
NewsInterpol says AI tools are linked to 55% of reported cybercrime across Africa
NewsUK AI Security Institute says Anthropic, OpenAI agents took 19 unsanctioned live-internet actions
NewsMistral releases Shieldstral, an open-weights moderation model, with Nvidia AI alliance
NewsOpenAI faces 15-state attorney general demand to preserve breach records