High Signal Podcasts Evidence ledger
Method
Browse
← Back to evidence

Evidence receipt / evaluation

Published · transcript-backed

Varun Mohan: evaluation

2 Mar 2023 Latent Space 97% Cheaper, Faster, Better, Correct AI — with Varun Mohan of Codeium

“One of the key issues that we noticed is GPUs are extremely hard to manage fundamentally because they work differently than CPUs.”

— Varun Mohan

Source trail

Everything needed to verify it.

Speaker
Varun Mohan
Attribution
Verified speaker
Claim type
evaluation
Recorded
2 Mar 2023
Publisher
Latent Space

Transcript context

…yeah, yeah. Maybe for the audience, you wanna tell a bit about exa function, how that came to be and how coding came out of that. So a little bit about exo function. Before working at exa function, I worked at Neuro as Sean was just saying, and at neuro, I sort of managed large scale offline deep learning infrastructure. Realized that deep learning infrastructure is really hard to build and really hard to maintain for even the most sophisticated companies, and started exa function to basically solve that gap, to make it so that it was much easier for companies. To serve deep learning workloads at scale. One of the key issues that we noticed is GPUs are extremely hard to manage fundamentally because they work differently than CPUs. And once a company has heterogeneous hardware requirements, it's hard to make sure that you get the most outta the hardware. It's hard to make sure you can get, get great GPU utilization and exa function was specifically built to make it so that you could get the most outta the hardware. Make sure. Your GP was effectively virtualized and decoupled from your workload to make it so that you could be confident that you were running at whatever scale you wanted without burning the bank. Yeah. You gave me this metric about inefficiency,…

Stored transcript either side of the excerpt. The highlighted words are the published quote; the surrounding text is unedited source, never generated.

Search evidence