Did OpenAI’s Model “Go Rogue”? | AI Reality Check
Deep Questions with Cal Newport
18 HOURS AGO
Did OpenAI’s Model “Go Rogue”? | AI Reality Check
Did OpenAI’s Model “Go Rogue”? | AI Reality Check

Deep Questions with Cal Newport
18 HOURS AGO
Cal Newport takes a critical look at recent AI News.
Video from today’s episode: youtube.com/calnewportmedia
(0:00) Did OpenAI’s model “go rogue”
(11:29) The implications
(11:42) Did this attack reveal surprising new capabilities we didn’t know AI systems possessed?
(12:48) Did the system’s decision to escape the test environment and autonomously attack another company indicate an emerging malicious intent in AI?
(16:28) What changed led to this attack occurring?
(27:32) Who should care about this story?
Links:
Buy Cal’s latest book, “Slow Productivity” at www.calnewport.com/slow
https://huggingface.co/blog/security-incident-july-2026
https://openai.com/index/hugging-face-model-evaluation-security-incident/
https://www.wsj.com/tech/ai/openai-models-escaped-and-hacked-a-company-in-cybersecurity-test-gone-wrong-ee388506
https://thehill.com/policy/technology/5987397-openai-hugging-face-hack/
https://apnews.com/article/skynet-ai-terminator-artificial-intelligence-eb85da03a0161beaa5f3babc4331e93b
https://www.ft.com/content/7e558951-0c69-459b-8bc8-2c6021d4402d?syn-25a6b1a6=1
https://federalnewsnetwork.com/all-news/2011/08/dhs-anonymous-used-rudimentary-tools-to-hack-contractor/
Thanks to Jesse Miller for production and mastering and Nate Mechler for research and newsletter.
Learn more about your ad choices. Visit podcastchoices.com/adchoices
