High Signal Podcasts Evidence ledger
Method
Browse
← Back to evidence

Evidence receipt / belief

Published · transcript-backed

Roman Yampolskiy: belief

2 Jun 2024 Lex Fridman Podcast #431 – Roman Yampolskiy: Dangers of Superintelligent AI

“The previous model we learned about after we finished training it, what it was capable of.”

— Roman Yampolskiy

Source trail

Everything needed to verify it.

Speaker
Roman Yampolskiy
Attribution
Verified speaker
Claim type
belief
Recorded
2 Jun 2024
Publisher
Lex Fridman Podcast

Transcript context

…Right. That happens when you go from the predictable to the unpredictable very quickly. But it’s not obvious to me that AI systems would gain capabilities so quickly that you won’t be able to collect enough data to study the benefits and risks. We’re literally doing it. The previous model we learned about after we finished training it, what it was capable of. Let’s say we stopped GPT-4 training run around human capability, hypothetically. We start training GPT- 5 and I have no knowledge of insider training runs or anything and started that point of about human and we train it for the next nine months. Maybe two months in, it becomes super intelligent. We continue training it. At the time when we start testing it, it is already a dangerous system. How dangerous? I have no idea, but never people training it. At the training stage, but then there’s a testing stage inside the company, they can start getting intuition about what the system is capable to do. You’re saying that somehow from leap from GPT-4 to GPT-5 can happen, the kind of leap where GPT-4 was controllable and GPT-5 is no longer controllable and we get no insights from using GPT-4 about the fact that GPT-5 will be uncontrollable. That’s the situation you’re concerned about. Where there leap from N, to N plus one will be such that an uncontrollable system is created without any ability for us to anticipate that.…

Stored transcript either side of the excerpt. The highlighted words are the published quote; the surrounding text is unedited source, never generated.

Search evidence