Ilya Sutskever — We're moving from the age of scaling to the age of research
Dwarkesh Podcast
2025/11/25
Ilya Sutskever — We're moving from the age of scaling to the age of research
Ilya Sutskever — We're moving from the age of scaling to the age of research

Dwarkesh Podcast
2025/11/25
Shownote
Shownote
Ilya & I discuss SSI’s strategy, the problems with pre-training, how to improve
the generalization of AI models, and how to ensure AGI goes well.
Watch on YouTube [https://youtu.be/aR20FWCCjAs]; read the transcript
[https://www.dwarkesh.com/p/ilya-sutskever-2].
Sponsors
* Gemini 3 [https://gemini.google] is the first model I’ve used that can find
connections I haven’t anticipated. I recently wrote a blog post on RL’s
information efficiency, and Gemini 3 helped me think it all through. It also
generated the relevant charts and ran toy ML experiments for me with zero bugs.
Try Gemini 3 today at gemini.google [https://gemini.google]
* Labelbox [https://labelbox.com/dwarkesh] helped me create a tool to transcribe
our episodes! I’ve struggled with transcription in the past because I don’t just
want verbatim transcripts, I want transcripts reworded to read like essays.
Labelbox helped me generate the exact data I needed for this. If you want to
learn how Labelbox can help you (or if you want to try out the transcriber tool
yourself), go to labelbox.com/dwarkesh [https://labelbox.com/dwarkesh]
* Sardine [https://sardine.ai/dwarkesh] is an AI risk management platform that
brings together thousands of device, behavior, and identity signals to help you
assess a user’s risk of fraud & abuse. Sardine also offers a suite of agents to
automate investigations so that as fraudsters use AI to scale their attacks, you
can use AI to scale your defenses. Learn more at sardine.ai/dwarkesh
[https://sardine.ai/dwarkesh]
To sponsor a future episode, visit dwarkesh.com/advertise
[https://www.dwarkesh.com/advertise].
Timestamps
(00:00:00) – Explaining model jaggedness
(00:09:39) - Emotions and value functions
(00:18:49) – What are we scaling?
(00:25:13) – Why humans generalize better than models
(00:35:45) – SSI’s plan to straight-shot superintelligence
(00:46:47) – SSI’s model will learn from deployment
(00:55:07) – How to think about powerful AGIs
(01:18:13) – “We are squarely an age of research company”
(01:20:23) – Self-play and multi-agent
(01:32:42) – Research taste
Get full access to Dwarkesh Podcast at www.dwarkesh.com/subscribe
[https://www.dwarkesh.com/subscribe?utm_medium=podcast&utm_campaign=CTA_4]
Highlights
Highlights
In this deep-dive conversation, Ilya Sutskever explores the frontiers of AI research, focusing on the limitations of current models and the path toward safe, general, and ultimately superintelligent systems. The discussion moves beyond benchmark performance to examine the foundational challenges in learning efficiency, generalization, and alignment.
Chapters
Chapters
Explaining model jaggedness
00:00Emotions and value functions
09:39What are we scaling?
18:49Why humans generalize better than models
25:13SSI’s plan to straight-shot superintelligence
35:45SSI’s model will learn from deployment
46:47How to think about powerful AGIs
55:07“We are squarely an age of research company”
1:18:13Self-play and multi-agent
1:30:26Research taste
1:32:42Transcript
Transcript
Ilya Sutskever: you know what's crazy, that all of this is real? Yeah, don't you think so? Like all this ai stuff and all this big era? Yeah, that it's happened, like, isn't it straight out of science fiction?
Dwarkesh Patel: Yeah, another thing that's cr...
