Evidence receipt / belief
Published · transcript-backedSoumith Chintala: belief
6 Mar 2024 Latent Space Open Source AI is AI we can Trust — with Soumith Chintala of Meta AI
“Like I think Together and Fireworks and all these people are trying to build some faster CUDA kernels and faster, you know, hardware kernels in general.”
Source trail
Everything needed to verify it.
- Speaker
- Soumith Chintala
- Attribution
- Verified speaker
- Claim type
- belief
- Recorded
- 6 Mar 2024
- Publisher
- Latent Space
Transcript context
…And he took it very well. And we'll have him on at some point and we'll discuss it. But I think it's important for, I think the market being maturing enough that people start caring and competing on these kinds of things means that we need to establish what best practice is because otherwise everyone's going to play dirty. Yeah, absolutely. My view of the LLM inference market in general is that it's the laundromat model. Like the margins are going to drive down towards the bare minimum. It's going to be all kinds of arbitrage between how much you can get the hardware for and then how much you sell the API and how much latency your customers are willing to let go. You need to figure out how to squeeze your margins. Like what is your unique thing here? Like I think Together and Fireworks and all these people are trying to build some faster CUDA kernels and faster, you know, hardware kernels in general. But those modes only last for a month or two. These ideas quickly propagate. Even if they're not published?…
Stored transcript either side of the excerpt. The highlighted words are the published quote; the surrounding text is unedited source, never generated.