Join Nostr
2026-09-01 07:32:53 UTC

Juliet E McKenna on Nostr: I have read a technical but comprehensible explanation of the OpenAI 'rogue software' ...

I have read a technical but comprehensible explanation of the OpenAI 'rogue software' event - with close attention as the detail relates to IT at the outer limit of my personal understanding.

Summing up is clear

"If you optimize a model to find exploits, you should expect it to find them — and prepare for that. OpenAI did not. They built a model, took the safeguards off, gave it the ExploitGym task, let it run, and didn't even monitor it. That's human decision-making."

https://mail.cyberneticforests.com/models-dont-go-rogue/