Evidence receipt / prediction
Published · transcript-backedDwarkesh Patel: prediction
6 Apr 2023 Dwarkesh Podcast Eliezer Yudkowsky — Why AI will kill us, aligning LLMs, nature of intelligence, SciFi, & rationality
“I expect them to be better than humans at science than they are at power seeking, because we had greater selection pressures for power seeking in our ancestral environment than we did for science.”
Source trail
Everything needed to verify it.
- Speaker
- Dwarkesh Patel
- Attribution
- Verified speaker
- Claim type
- prediction
- Recorded
- 6 Apr 2023
- Publisher
- Dwarkesh Podcast
Transcript context
…No. You have spoken it. It exists. It cannot be called back. There are no take backs. There is no going back. There is no going back. Go ahead. Okay, so here’s another story. I expect them to be better than humans at science than they are at power seeking, because we had greater selection pressures for power seeking in our ancestral environment than we did for science. And while at a certain point both of them come along as a package, maybe they can be at varying levels, so you have this sort of early model that is kind of human-level, except a little bit ahead of us in science. You ask it to help us align the next version of it, then the next version of it is more aligned because we have its help and sort of like this inductive thing where the next version helps us align the version. Where do people have this notion of getting AIs to help you do your AI alignment homework? Why can we not talk about having it enhance humans instead?…
Stored transcript either side of the excerpt. The highlighted words are the published quote; the surrounding text is unedited source, never generated.