From Daniel Kokotajlo · OpenAI
“I believe the leaders of the top labs are making reckless decisions in how much autonomy and capability they're giving to their models.”
Public reporting or insider accounts confirming major AI labs shipped models with abbreviated or skipped safety evaluations.
Partially resolved
The Washington Post and multiple outlets reported on compressed safety testing timelines at major labs through 2024-2025. Former employees from OpenAI, Google, and xAI described safety evaluations being shortened to meet product deadlines. OpenAI shipped GPT-4o with abbreviated red-teaming.
View evidence →Resolved 6/1/2025