The pace of AI model releases is accelerating, forcing safety evaluators to shift from pre-deployment to post-deployment research.
The case
Ryan Greenblatt predicts that AI misalignment will become really concerning in about three years from now.
“By my default modal timeline, I think shit is really, really crazy and concerning from a misalignment perspective more like three years from now.”Ryan Greenblatt · 11 Aug 2026
UK AISI is shifting from pre-deployment evaluations to longer post-deployment research collaborations due to the increasing pace of model releases.
“We are shifting a lot of that work not least because the pace of model releases is increasing to longer research collaborations that might go back either are post deployment or they go back further before deployment is finalized.”Geoffrey Irving · 1 Mar 2026
The pushback
Anthropic's Mythos model was available internally to Anthropic employees in February but only released to the public in June.
“We saw, for example, that Mythos was available internally to Anthropic employees in February, but only released to the public in, I think, June, actually.”Ryan Greenblatt · 11 Aug 2026
Topics
Signal Headquarters · compiled from attributed public discussion. Last updated 2026-08-11.