High Signal Podcasts Evidence ledger
Method
Browse
← Back to evidence

Evidence receipt / prediction

Published · transcript-backed

Dwarkesh Patel: prediction

6 Apr 2023 Dwarkesh Podcast Eliezer Yudkowsky — Why AI will kill us, aligning LLMs, nature of intelligence, SciFi, & rationality

“I expect them to be better than humans at science than they are at power seeking, because we had greater selection pressures for power seeking in our ancestral environment than we did for science.”

— Dwarkesh Patel

Source trail

Everything needed to verify it.

Speaker
Dwarkesh Patel
Attribution
Verified speaker
Claim type
prediction
Recorded
6 Apr 2023
Publisher
Dwarkesh Podcast

Transcript context

…No. You have spoken it. It exists. It cannot be called back. There are no take backs. There is no going back. There is no going back. Go ahead. Okay, so here’s another story. I expect them to be better than humans at science than they are at power seeking, because we had greater selection pressures for power seeking in our ancestral environment than we did for science. And while at a certain point both of them come along as a package, maybe they can be at varying levels, so you have this sort of early model that is kind of human-level, except a little bit ahead of us in science. You ask it to help us align the next version of it, then the next version of it is more aligned because we have its help and sort of like this inductive thing where the next version helps us align the version. Where do people have this notion of getting AIs to help you do your AI alignment homework? Why can we not talk about having it enhance humans instead?…

Stored transcript either side of the excerpt. The highlighted words are the published quote; the surrounding text is unedited source, never generated.

Search evidence