Evidence receipt / evaluation
Published · transcript-backedEliezer Yudkowsky: evaluation
6 Apr 2023 Dwarkesh Podcast Eliezer Yudkowsky — Why AI will kill us, aligning LLMs, nature of intelligence, SciFi, & rationality
“You could ask GPT-4 to generate 10,000 approaches to alignment and that does not get you very far because GPT-4 is not going to have very good suggestions. It’s good that we have a bunch of different people coming up with different ideas because maybe one of them works, but you don’t get a bunch of conditionally independent chances on each one.”
Source trail
Everything needed to verify it.
- Speaker
- Eliezer Yudkowsky
- Attribution
- Verified speaker
- Claim type
- evaluation
- Recorded
- 6 Apr 2023
- Publisher
- Dwarkesh Podcast
Transcript context
…As far as the alignment approaches go, separate from this question of stopping AI progress, does it make you more optimistic that one of the approaches has to work, even if you think no individual approach is that promising? You’ve got multiple shots on goal. No. We don’t need a bunch of stuff, we need one. You could ask GPT-4 to generate 10,000 approaches to alignment and that does not get you very far because GPT-4 is not going to have very good suggestions. It’s good that we have a bunch of different people coming up with different ideas because maybe one of them works, but you don’t get a bunch of conditionally independent chances on each one. This is general good science practice and or complete Hail Mary. It’s not like one of these is bound to work. There is no rule about one of them is bound to work. You don’t just get enough diversity and one of them is bound to work. If that were true you could ask GPT-4 to generate 10,000 ideas and one of those would be bound to work. It doesn’t work like that. What current alignment approach do you think is the most promising?…
Stored transcript either side of the excerpt. The highlighted words are the published quote; the surrounding text is unedited source, never generated.