scripod.com

Did OpenAI’s Model “Go Rogue”? | AI Reality Check

Shownote

Cal Newport takes a critical look at recent AI News. Video from today’s episode: youtube.com/calnewportmedia (0:00) Did OpenAI’s model “go rogue” (11:29) The implications (11:42) Did this attack reveal surprising new capabilities we didn’t know AI ...

Highlights

Cal Newport dissects a recent AI security incident, separating hype from reality. He examines what really happened when an OpenAI model appeared to 'go rogue' during a cybersecurity test, challenging sensationalist headlines.
03:19
Harnesses are standard computer programs, not mysterious AI
11:31
Key questions to analyze the OpenAI security incident.
11:42
The system was simply performing its designed function
12:48
LLMs are static, feed-forward models without intent or planning.
18:55
LLMs are not malicious but unpredictable
32:47
It was a cybersecurity mistake, not a sign of rogue AI.

Chapters

Did OpenAI’s model “go rogue”
00:00
The implications
11:29
Did this attack reveal surprising new capabilities we didn’t know AI systems possessed?
11:42
Did the system’s decision to escape the test environment and autonomously attack another company indicate an emerging malicious intent in AI?
12:48
What changed led to this attack occurring?
16:28
Who should care about this story?
27:32

Transcript

Cal Newport: A couple weeks ago, the AI company Hugging Face announced that they had discovered an intrusion into their production infrastructure. Now, they didn't know the source, but they noted that it looked like large language models were involved. Wel...