High Signal Podcasts Evidence ledger
Method
Browse
← Back to evidence

Evidence receipt / belief

Published · transcript-backed

Sundar Pichai: belief

7 Apr 2026 Cheeky Pint The history and future of AI at Google, with Sundar Pichai

“Depending on what we think you're doing, some people may get a latency budget of 30 milliseconds or 10 milliseconds.”

— Sundar Pichai

Source trail

Everything needed to verify it.

Speaker
Sundar Pichai
Attribution
Verified speaker
Claim type
belief
Recorded
7 Apr 2026
Publisher
Cheeky Pint

Transcript context

…As I think about the AI race in 2026, one thing that strikes me is Google has for so long had speed as the place it tries to differentiate. Original Google Search was really fast and famously displayed the search query time within the results, sort of showing off. Then Gmail Fast Search compared to the competitors of the time, or Chrome, compared to the competitors of the time. And now, I use all of the AI services for different things, but Gemini on TPUs is just so fast. I'm curious how much this is part of the explicit product strategy and how you think of it, or it's much more nuanced than that. I've always internalized speed. Let's call it latency for this purpose, and as one of the distinguishing features of a great product. Also, it almost always reflects the technical underpinnings of the product having been done well. There's a different speed which matters, too, which is the speed of shipping and iteration and release cycles. Both are important. But you talk about latency. It's easy to say you want latency, but you're constantly adding capabilities. The capability frontier is progressing. There's some sense of, how do you balance that? That's where it gets more complicated. But to give an example, like Search, I was speaking with the teams. They now have for sub-teams, latency budgets in the milliseconds. You get 50% credit... If you ship something which shaves off 3 milliseconds, you earn 1. 5 milliseconds for your latency budget, and 1.5 milliseconds gets passed on to the user. Depending on what we think you're doing, some people may get a latency budget of 30 milliseconds or 10 milliseconds. You can use it, but you have rigorous reviews against that. But that's how much we think it matters. For context, I guess humans pick it up in the low hundreds of milliseconds, is that correct? In terms of where it actually impacts?…

Stored transcript either side of the excerpt. The highlighted words are the published quote; the surrounding text is unedited source, never generated.

Search evidence