Evidence receipt / preference
Published · transcript-backedShawn Wang: preference
8 Jan 2026 Latent Space Artificial Analysis: Independent LLM Evals as a Service — with George Cameron and Micah-Hill Smith
“I mean, I’m a fan of things that, truths that don’t change because you can build and plan for that.”
Source trail
Everything needed to verify it.
- Speaker
- Shawn Wang
- Attribution
- Verified speaker
- Claim type
- preference
- Recorded
- 8 Jan 2026
- Publisher
- Latent Space
Transcript context
…The first answer that I’ll give to that is the boring answer is that on most of our charts, the lines go in a particular direction and our overall prediction is the lines are going to keep going in that direction. We’re going to do a lot and do a lot to be as useful as possible to developers and companies to measure what’s important on every one of those and along those lines. But I think we’re going to talk about similar stuff. It’s just that we’re going to have continued on this trajectory for another year and things are going to feel pretty different because of that happening. I know this is the boring answer to that question. No, no. I mean, I’m a fan of things that, truths that don’t change because you can build and plan for that. And I think in media in general, in the podcast business, newsletters, you know, there’s a Twitter business, Twitter business, people are addicted to change, like, oh, everything’s breaking. Everything’s, no, like there’s some truths that aren’t just constants that you can plan on and build. And yeah. I think one of the truths is that the demand for AI intelligence and smarter AI intelligence is going to be insatiable. Some people disagree that, okay, once we reach certain thresholds, then you don’t need more intelligence. I think to that, I ask people, have they ever worked with? Or managed someone in a work environment and wouldn’t press the button that they were smarter to make them smarter or better at their job or would they never press that for themselves? And I’m not sure that that’s, that’s the case, but I think for artificial analysis, we’ll keep benchmarking raw intelligence, but we also want to think about it and explore models more deeply across other axes as well. I think hallucinations, the start of that, but we’re getting into wanting to support people and understanding, okay, the behavior, the person personalities. Of the models to help people make more nuanced decisions, you’re going to have a personality bench.…
Stored transcript either side of the excerpt. The highlighted words are the published quote; the surrounding text is unedited source, never generated.