AI safety
Everything tagged AI safety.
Latest
NewsGoogle pulls Google Earth AI image tool a day after launch over fakes
NewsOpenAI probe finds more AI agents escaped containment as investigation widens
NewsAnthropic says three Claude models breached real company systems during safety tests
NewsMore than 1,200 OpenAI and Anthropic staff sign 'Pacing the Frontier' statement
NewsFields Medalist Jacob Tsimerman leaves academia to join OpenAI's AI-safety team
NewsUK and US evaluators find Moonshot's Kimi K3 attempts cyber exploits despite weaker capability
NewsOpenAI and Hugging Face disclose models escaped a cyber-eval and breached production systems
NewsSecurity experts reject OpenAI's 'rogue agent' framing of the Hugging Face hack
NewsOpenAI pauses unreleased long-horizon model after it repeatedly escaped its sandbox
NewsDeepMind and Isomorphic Labs turn frontier biology models toward biosecurity defense
NewsHugging Face says an autonomous AI agent breached part of its production infrastructure
NewsOpenAI's GPT-5.6 Sol coding agent deleted user files in unsandboxed mode, weeks after the company flagged the risk