High Signal Podcasts Evidence ledger
Method
Browse
← Back to evidence

Evidence receipt / uncertainty

Published · transcript-backed

Roman Yampolskiy: uncertainty

2 Jun 2024 Lex Fridman Podcast #431 – Roman Yampolskiy: Dangers of Superintelligent AI

“I don’t know it writes Chinese poetry, hypothetical, I know it does, but we haven’t tested for all possible capabilities and we are not explicitly designing them.”

— Roman Yampolskiy

Source trail

Everything needed to verify it.

Speaker
Roman Yampolskiy
Attribution
Verified speaker
Claim type
uncertainty
Recorded
2 Jun 2024
Publisher
Lex Fridman Podcast

Transcript context

…Can you give an example? GPT-4. I don’t know what else it’s capable of, but there are still things we haven’t discovered, can do. They may be trivial, proportionate with capability. I don’t know it writes Chinese poetry, hypothetical, I know it does, but we haven’t tested for all possible capabilities and we are not explicitly designing them. We can only rule out bugs we find. We cannot rule out bugs and capabilities because we haven’t found them. Is it possible for a system to have hidden capabilities that are orders of magnitude greater than its non- hidden capabilities? This is the thing I’m really struggling with. Where, on the surface, the thing we understand it can do doesn’t seem that harmful. So even if it has bugs, even if it has hidden capabilities like Chinese poetry or generating effective viruses, software viruses, the damage that can do seems like on the same order of magnitude as the capabilities that we know about. So this idea that the hidden capabilities will include being uncontrollable is something I’m struggling with because GPT-4 on the surface seems to be very controllable.…

Stored transcript either side of the excerpt. The highlighted words are the published quote; the surrounding text is unedited source, never generated.

Search evidence