High Signal Podcasts Evidence ledger
Method
Browse
← Back to evidence

Evidence receipt / uncertainty

Published · transcript-backed

Mark Zuckerberg: uncertainty

18 Apr 2024 Dwarkesh Podcast Mark Zuckerberg — Llama 3, $10B models, Caesar Augustus, & 1 GW datacenters

“I don't know what that ratio is going to be but I consider the generation of synthetic data to be more inference than training today.”

— Mark Zuckerberg

Source trail

Everything needed to verify it.

Speaker
Mark Zuckerberg
Attribution
Verified speaker
Claim type
uncertainty
Recorded
18 Apr 2024
Publisher
Dwarkesh Podcast

Transcript context

…But it doesn’t have to be in the same place, right? If distributed training works, it can be distributed. Well, I think that is a big question, how that's going to work. It seems quite possible that in the future, more of what we call training for these big models is actually more along the lines of inference generating synthetic data to then go feed into the model. I don't know what that ratio is going to be but I consider the generation of synthetic data to be more inference than training today. Obviously if you're doing it in order to train a model, it's part of the broader training process. So that's an open question, the balance of that and how that plays out. Would that potentially also be the case with Llama-3, and maybe Llama-4 onwards? As in, you put this out and if somebody has a ton of compute, then they can just keep making these things arbitrarily smarter using the models that you've put out. Let’s say there’s some random country, like Kuwait or the UAE, that has a ton of compute and they can actually just use Llama-4 to make something much smarter.…

Stored transcript either side of the excerpt. The highlighted words are the published quote; the surrounding text is unedited source, never generated.

Search evidence