High Signal Podcasts Evidence ledger
Method
Browse
← Back to evidence

Evidence receipt / commitment

Published · transcript-backed

Nicholas Carlini: commitment

29 Aug 2024 Latent Space Why you should write your own LLM benchmarks — with Nicholas Carlini, Google DeepMind

“I will take whatever tools are available in front of me and try and see if I can use them in meaningful ways.”

— Nicholas Carlini

Source trail

Everything needed to verify it.

Speaker
Nicholas Carlini
Attribution
Verified speaker
Claim type
commitment
Recorded
29 Aug 2024
Publisher
Latent Space

Transcript context

…No, I canceled my OpenAI subscription, so I'm a Claude boy. Do you have a way to think about this like one-offs software thing? One way I talk to people about it is like LLMs are kind of converging to like semantic serverless functions, you know, like you can say something and like it can run the function in a way and then that's it. It just kind of dies there. Do you have a mental model to just think about how long it should live for and like anything like that? I don't think I have anything interesting to say here, no. I will take whatever tools are available in front of me and try and see if I can use them in meaningful ways. And if they're helpful, then great. If they're not, then fine. And like, you know, there are lots of people that I'm very excited about seeing all these people who are trying to make better applications that use these or all these kinds of things. And I think that's amazing. I would like to see more of it, but I do not spend my time thinking about how to make this any better. What's the most underrated thing in the list? I know there's like simplified code, solving boring tasks, or maybe is there something that you forgot to add that you want to throw in there?…

Stored transcript either side of the excerpt. The highlighted words are the published quote; the surrounding text is unedited source, never generated.

Search evidence