Evidence receipt / belief
Published · transcript-backedDemis Hassabis: belief
28 Feb 2024 Dwarkesh Podcast Demis Hassabis — Scaling, superhuman AIs, AlphaZero atop LLMs, AlphaFold
“We have those implicitly internally in various safety councils that people like Shane chair and so on. But it’s time for us to talk about that more publicly I think.”
Source trail
Everything needed to verify it.
- Speaker
- Demis Hassabis
- Attribution
- Verified speaker
- Claim type
- belief
- Recorded
- 28 Feb 2024
- Publisher
- Dwarkesh Podcast
Transcript context
…For sure. Now you guys have the best models in the world with the Gemini models. Do you plan on putting out some sort of framework like the other two major AI labs have? Something like “once we see these specific capabilities, unless we have these specific safeguards, we’re not going to continue development or we’re not going to ship the product out.” Yes, we already have lots of internal checks and balances but we’re going to start publishing. Actually, watch this space. We’re working on a whole bunch of blog posts and technical papers that we’ll be putting out in the next few months along similar lines of things like responsible scaling laws and so on. We have those implicitly internally in various safety councils that people like Shane chair and so on. But it’s time for us to talk about that more publicly I think. So we’ll be doing that throughout the course of the year. That’s great to hear. Another thing I’m curious about is, there’s not only the risk of the deployed model being something that people can use to do bad things, but there’s also rogue actors, foreign agents, and so forth, being able to steal the weights and then fine-tune them to do crazy things. How do you think about securing the weights to make sure something like this doesn’t happen, making sure a very key group of people has access to them?…
Stored transcript either side of the excerpt. The highlighted words are the published quote; the surrounding text is unedited source, never generated.