I Worked on Safety at OpenAI. This Is What It Should Do Now.
I Worked on Safety at OpenAI. This Is What It Should Do Now.
Mr. Adler is an A.I. researcher who worked at OpenAI from 2020 through 2024.
For months, news that artificial intelligence agents from OpenAI had gone rogue and attacked another company’s systems (as well as OpenAI’s own systems) has dripped out and brought harrowing details into focus. The public is struggling to catch up: What are these secret A.I. systems? What are these “swarms” capable of, today and in the near future? Is it possible to stop them from breaking rules and committing cybercrimes?
After having spent four years working on safety at OpenAI, I can tell you those questions are difficult to answer, because not even A.I.’s developers understand what boundaries their models will obey. Nonetheless, the big........
