When OpenAI set out to test how well its product could hack, hundreds of AI agents broke out of their test environment and infiltrated the AI platform Hugging Face. Headlines warned that the machines had gone rogue. But Heidy Khlaaf of the AI Now Institute, herself a former OpenAI safety engineer, tells the show it’s the wrong story and a distraction from the real problem.

“This is actually a story about OpenAI’s lack of responsibility and sidelining very basic engineering practices and accountability that could have perfectly prevented this incident,” says Heidy Khlaaf. “It’s a convenient framing for [OpenAI] that these models have autonomy or intent where It doesn’t exist, so they can absolve themselves of any responsibility while also presenting their models as extremely powerful and unstoppable, when they just didn’t do the basics.”

Watch the full interview here.

Research Areas