Evidence receipt / commitment
Published · transcript-backedCat Wu: commitment
23 Apr 2026 Lenny's Podcast How Anthropic’s product team moves faster than anyone else | Cat Wu (Head of Product, Claude Code)
“we tried to build a code review product a few times and we've launched like simple versions of code review which is the slash code review command in the past and it was only with the most recent models that we felt like okay this code review is so good that our engineering team relies on this code review to pass before we merge prs and we found that this was We've always dreamed of Claude being able to be a reliable code reviewer that we can confidently feel catches the majority of bugs.”
— Cat Wu
Source trail
Everything needed to verify it.
- Speaker
- Cat Wu
- Attribution
- Verified speaker
- Claim type
- commitment
- Recorded
- 23 Apr 2026
- Publisher
- Lenny's Podcast
Transcript context
…I forget who said this on the podcast that the model will eat your harness for breakfast. And what I'm hearing here, essentially, you you remove things over time that you've had to add on top of the model where it was not operating the way you want it. And essentially, as the models get smarter, you just it becomes simpler and simpler for it's just to do the thing you want it to do. Yeah, we can remove a lot of prompting interventions every time the model gets smarter. And we actually do this every time we launch a model. We read through the entire system prompt and we reflect on, okay, for each of these sections, does the model really need this reminder anymore? And if not, we'll remove it. The most exciting thing that new models unlocks though, it's just like entirely new features. So there's all the features that we've been testing out with prior models and the accuracy wasn't high enough for us to want to launch them. And so one example of this is code review. we tried to build a code review product a few times and we've launched like simple versions of code review which is the slash code review command in the past and it was only with the most recent models that we felt like okay this code review is so good that our engineering team relies on this code review to pass before we merge prs and we found that this was We've always dreamed of Claude being able to be a reliable code reviewer that we can confidently feel catches the majority of bugs. And it was only with like Opus 4.5 and 4.6 and Sonic 4.6 that we felt like, okay, we are now able to run multiple code review agents simultaneously to traverse the entirety of the code base and to synthesize a set of like real issues that an engineer needs to address before merge and so this is like a new capability that the the newest models have unlocked this is another trend that is very common on this podcast of build something that will possibly be possible in the next six months be kind of at the edge of what's working sort of and then it'll catch up and then it'll be an amazing product and you'll be ahead of everyone…
Stored transcript either side of the excerpt. The highlighted words are the published quote; the surrounding text is unedited source, never generated.