Sveriges mest populära poddar
LessWrong (30+ Karma)
LessWrong (30+ Karma)

“OpenAI President Brockman says HuggingFace incident model had not been alignment-trained” by Caspar Oesterheld

1 min•15 september 2026

Om avsnittet

On today's episode of the podcast "Odd Lots", OpenAI President Greg Brockman said (at around 8:40): "This model that did/had the HuggingFace incident actually had not gone through our alignment training, yet." I assume Brockman is specifically referring to the "Highly Persistent Internal Model" as it's called in the METR/Redwood report. As far as I know, OpenAI has not said before whether this model had been alignment-trained or not.

---

First published:
September 14th, 2026

Source:
https://www.lesswrong.com/posts/67gHvbmFeacXi2jCZ/openai-president-brockman-says-huggingface-incident-model

---

Narrated by TYPE III AUDIO.

LessWrong (30+ Karma) med LessWrong finns tillgänglig på flera plattformar. Informationen på denna sida kommer från offentliga podd-flöden.