Evidence receipt / uncertainty
Published · transcript-backedTrenton Bricken: uncertainty
22 May 2025 Dwarkesh Podcast Is RL + LLMs enough for AGI? — Sholto Douglas & Trenton Bricken
“” For background context, Nicholas Carlini is a researcher who actually was at DeepMind and has now come over to Anthropic. But the model says, "Oh, I don't know who that is.”
Source trail
Everything needed to verify it.
- Speaker
- Trenton Bricken
- Attribution
- Verified speaker
- Claim type
- uncertainty
- Recorded
- 22 May 2025
- Publisher
- Dwarkesh Podcast
Transcript context
…Yeah. That's so funny. Or Transluce has another example of this, where you ask a Llama model, “who is Nicholas Carlini? ” For background context, Nicholas Carlini is a researcher who actually was at DeepMind and has now come over to Anthropic. But the model says, "Oh, I don't know who that is. I couldn't possibly speculate." But if you look at the features behind the scenes, you see a bunch light up for AI, computer security, all the things that Nicholas Carlini does. Interpretability becomes dramatically more important as you shift in this direction of Neuralese.…
Stored transcript either side of the excerpt. The highlighted words are the published quote; the surrounding text is unedited source, never generated.