policy-safety-and-incidents
Everything tagged policy-safety-and-incidents.
Latest
NewsHacktron says Claude helped breach OpenAI employee accounts
NewsAnthropic Opens Life Sciences Verification Program Beta
NewsOpenAI publishes misalignment disclosure framework and six incident reports
NewsOpenAI says it is in AI safety talks with Anthropic and Google DeepMind
NewsOpenAI reportedly asked Congress about a coordinated AI slowdown
NewsAnthropic says it disrupted Claude misuse across cyberattacks and weapons research
NewsAnthropic commits to embedded external evaluators as Amodei calls for slower AI gains
NewsResearchers link May RubyGems package flood to OpenAI agents
NewsOpenAI appoints Paul Christiano to Foundation Board and Safety and Security Committee
NewsOpenAI commits $1B in Daybreak access for essential-service cyber defenders
NewsOpenAI commits up to $5 million to independent research on AI and teen development
NewsOpenAI acknowledges German wiki agent incident, calls for wider disclosures