High Signal Podcasts Evidence ledger
Method
Browse
← Back to evidence

Evidence receipt / recommendation

Published · transcript-backed

Mentions personal use of Claude Max. Mentions personal use of Codex Pro.

17 Jun 2026 The Cognitive Revolution Radically Better Reasoning: Elicit's Andreas Stuhlmüller & Jungwon Byun on World Models for Research

“And certainly if you were to triple it from there, you'd be getting into something on the order of magnitude of parity with human headcount. What are you doing with it all too? Because I use my $200 Claude Max and my Codex Pro and I honestly don't even hit my limits that often.”

— Nathan Labenz

Source trail

Everything needed to verify it.

Speaker
Nathan Labenz
Attribution
Verified speaker
Claim type
recommendation
Recorded
17 Jun 2026
Publisher
The Cognitive Revolution

Transcript context

…I think even internally at Elicit, I'm not sure how many more multiples of token costs we can easily spend. So maybe taking myself as an example, I spend maybe $2,000 or so per week on tokens. And it could, maybe I could like double it or triple it, I don't know. But not much more than that for sure. So I do expect, and that is already influencing my behavior to some extent. So I don't actually currently use fast mode for these models for that reason because I don't feel the modular returns are high enough for most tasks. So I expect It's unlikely that I'll be like, oh, I need to switch over everything I do to and it's probably not going to happen. And both for my own usage and also this is already actually the case in the list of the app, I think more of a look like if like one smart orchestrator agent that then you know spins off like many other agents that have to do simpler tasks that just don't need to be the largest model. And I expect that will just become increasingly important, this sort of dispatching to a model of the right size so that you get the You get the intelligence when you need it, but you're not like just like multiplying your whole spend by some number that was an inefficient use of compute to begin with. So $2,000 is not a small amount. Is that a outlier with, does that make you an outlier within the company or is everybody doing that? If so, that would put your token cost like not presumably at the level of payroll because I assume you're paying your engineers more than that. But it would be like at least a not insignificant share. And certainly if you were to triple it from there, you'd be getting into something on the order of magnitude of parity with human headcount. What are you doing with it all too? Because I use my $200 Claude Max and my Codex Pro and I honestly don't even hit my limits that often. Now this may be API, which might be 10 times more. And so that could be a big part of it, but I sometimes feel a little ashamed that I'm like not redlining the account more than I am. What do you, how would you advise me maybe to, or what sort of personal bitter lessons have you learned that you're really like finding that token maxing is worth it? Yeah, I think I'm not sure I'm the top user of tokens at Elicit, but I'm probably at least in the top five. So I'm probably a little bit of an outlier. Second, yeah, I'm using the API. I could probably like save more money by being more clever about how to use various like pro accounts and stuff. I do have a fairly elaborate like system built on like Pi that like orchestrates between the different agents and uses like ChatGPT to double check Claude to then sometimes call Gemini to get like another take. And that is a little bit easier to do if you're on the API than if you're on the normal end user plans. That might just be part of the explanation here.…

Stored transcript either side of the excerpt. The highlighted words are the published quote; the surrounding text is unedited source, never generated.

Named in this claim

Books, apps, tools, and people.

Search evidence