“OpenAI President Brockman says HuggingFace incident model had not been alignment-trained” by Caspar Oesterheld
Podcast:LessWrong (30+ Karma) Published On: Tue Sep 15 2026 Description: On today's episode of the podcast "Odd Lots", OpenAI President Greg Brockman said (at around 8:40): "This model that did/had the HuggingFace incident actually had not gone through our alignment training, yet." I assume Brockman is specifically referring to the "Highly Persistent Internal Model" as it's called in the METR/Redwood report. As far as I know, OpenAI has not said before whether this model had been alignment-trained or not. --- First published: September 14th, 2026 Source: https://www.lesswrong.com/posts/67gHvbmFeacXi2jCZ/openai-president-brockman-says-huggingface-incident-model --- Narrated by TYPE III AUDIO.