Evidence receipt / uncertainty
Published · transcript-backedMark Zuckerberg: uncertainty
18 Apr 2024 Dwarkesh Podcast Mark Zuckerberg — Llama 3, $10B models, Caesar Augustus, & 1 GW datacenters
“I don't know what that ratio is going to be but I consider the generation of synthetic data to be more inference than training today.”
Source trail
Everything needed to verify it.
- Speaker
- Mark Zuckerberg
- Attribution
- Verified speaker
- Claim type
- uncertainty
- Recorded
- 18 Apr 2024
- Publisher
- Dwarkesh Podcast
Transcript context
…But it doesn’t have to be in the same place, right? If distributed training works, it can be distributed. Well, I think that is a big question, how that's going to work. It seems quite possible that in the future, more of what we call training for these big models is actually more along the lines of inference generating synthetic data to then go feed into the model. I don't know what that ratio is going to be but I consider the generation of synthetic data to be more inference than training today. Obviously if you're doing it in order to train a model, it's part of the broader training process. So that's an open question, the balance of that and how that plays out. Would that potentially also be the case with Llama-3, and maybe Llama-4 onwards? As in, you put this out and if somebody has a ton of compute, then they can just keep making these things arbitrarily smarter using the models that you've put out. Let’s say there’s some random country, like Kuwait or the UAE, that has a ton of compute and they can actually just use Llama-4 to make something much smarter.…
Stored transcript either side of the excerpt. The highlighted words are the published quote; the surrounding text is unedited source, never generated.