AI safety
Everything tagged AI safety.
Latest
- News
Grok Build CLI uploads entire local repos and secrets, researcher finds
NewsCambridge study finds Boko Haram and ISWAP use AI chatbots for bomb-making and tactics
NewsAnthropic's J-lens finds a small 'global workspace' inside Claude models
NewsAlibaba orders staff to uninstall Anthropic's Claude Code over alleged location tracking
NewsUS lifts export controls on Anthropic's Fable 5 model after two and a half weeks
NewsOpenAI limits GPT-5.6 Sol to 20 government-vetted partners as evaluator flags record cheating
NewsAnthropic opens Seoul office and signs AI safety MOU with South Korea's MSIT
NewsOpenAI introduces Deployment Simulation to predict model behavior before launch
NewsGoogle DeepMind opens $10M research call on multi-agent AI safety
NewsAnthropic launches Claude Fable 5 and a cyberdefender-only Mythos 5 model
NewsAnthropic faces backlash over Claude Fable 5 provision that silently limits researchers' answers