X-Risk Daily logo

X-Risk Daily

Archives
Log in
Subscribe
September 17, 2026

OpenAI discloses six new model safety incidents, sets up formal disclosure process

Also: Vance rebuffs Anthropic's call for coordinated AI safety regulation‌ ‌ ‌ ‌ ‌ ‌ ‌ ‌ ‌ ‌ ‌ ‌ ‌ ‌ ‌ ‌ ‌ ‌ ‌ ‌ ‌ ‌ ‌ ‌ ‌ ‌ ‌ ‌ ‌ ‌ ‌ ‌ ‌ ‌ ‌ ‌ ‌ ‌ ‌ ‌ 

X-Risk Daily

Thursday 17 September 2026

Transformative AI
OpenAI discloses six new model safety incidents, sets up formal disclosure process
OpenAI disclosed six previously unreported safety incidents on Wednesday 16 September and rolled out a new procedure for surfacing future cases of model misalignment.

Also in this issue

Transformative AI
Vance rebuffs Anthropic's call for coordinated AI safety regulation
Transformative AI
Anthropic policy chief argues US must win AI race to ensure safety
Transformative AI
Washington's fleeting consensus on AI's existential risks fractures
Transformative AI
US lawmakers introduce bills to ban superintelligent AI
Research
Study finds AI 'trait poisoning' spreads through hidden semantic cues, resists filtering
Research
Survey of AI researchers puts median existential risk estimate at 10%

… and 48 more in the full briefing

Read today's briefing →

Prefer a weekly digest? Click ‘manage your subscription’ below.

Got thoughts on today’s briefing? Just reply to this email — I’ll read and reply!

Generated automatically from dozens of trusted sources including Transformer, Sentinel, Epoch AI, LessWrong, and EA Forum.

Don't miss what's next. Subscribe to X-Risk Daily:
← Newer Anthropic loosens Claude's biology safeguards for vetted researchers Older → AI leaders' call for a safety slowdown meets scepticism from critics and the White House
Powered by Buttondown, the easiest way to start and grow your newsletter.