Evidence receipt / belief
Published · transcript-backedDemis Hassabis: belief
28 Feb 2024 Dwarkesh Podcast Demis Hassabis — Scaling, superhuman AIs, AlphaZero atop LLMs, AlphaFold
“The systems that are around today are not dangerous, in my opinion, but in a few years they might have potential.”
Source trail
Everything needed to verify it.
- Speaker
- Demis Hassabis
- Attribution
- Verified speaker
- Claim type
- belief
- Recorded
- 28 Feb 2024
- Publisher
- Dwarkesh Podcast
Transcript context
…Got it. Do you have a sense of what the converse answer would be? So what would have to be true where tomorrow morning you’re like “oh, man, I didn’t anticipate this.” You see some specific observation tomorrow morning that makes you say “we got to stop Gemini 2 training.” I could imagine that. This is where things like the sandbox simulations are important. I would hope we’re experimenting in a safe, secure environment when something very unexpected happens. There’s a new unexpected capability or something that we didn’t want. We explicitly told the system we didn’t want it but then it did and it lied about it. These are the kinds of things where one would want to then dig in carefully. The systems that are around today are not dangerous, in my opinion, but in a few years they might have potential. Then you would ideally pause and really get to the bottom of why it was doing those things before one continued. Going back to Gemini, I’m curious what the bottlenecks were in the development. Why not immediately make it one order of magnitude bigger if scaling works?…
Stored transcript either side of the excerpt. The highlighted words are the published quote; the surrounding text is unedited source, never generated.