Evidence receipt / prediction
Published · transcript-backedAman Sanger: prediction
22 Aug 2023 Latent Space Cursor.so: The AI-first Code Editor — with Aman Sanger of Anysphere
“I think I have some thoughts on this because there's the whole thing with chinchilla scaling and then people are now saying, oh, chinchilla scaling doesn't matter because of inference.”
Source trail
Everything needed to verify it.
- Speaker
- Aman Sanger
- Attribution
- Verified speaker
- Claim type
- prediction
- Recorded
- 22 Aug 2023
- Publisher
- Latent Space
Transcript context
…One of the reasons I harp on this is one of our pet themes is tracking the dataset to parameter ratio. And Copilot cannot be that big because it returns relatively quickly. So it's going to be in the low billions, right? So how do you do trillions of tokens to the low billions? That's interesting. Yeah. I think I have some thoughts on this because there's the whole thing with chinchilla scaling and then people are now saying, oh, chinchilla scaling doesn't matter because of inference. But Copilot could be a mixture of experts. That's one other speculation. I don't know if that's true. I mean, it probably wasn't the case at least a year or two ago. My guess is it's probably a small model that's very over-trained. From what I've heard, there are also lots of tricks you can do with caching where even if the model is quite big, it doesn't take, it effectively takes no time to ingest the entire prompt. Yeah. Semantic caching is what they are calling it, right? I guess if it roughly embeds to the same thing, just return the same thing.…
Stored transcript either side of the excerpt. The highlighted words are the published quote; the surrounding text is unedited source, never generated.