Evidence receipt / belief
Published · transcript-backedNathan Labenz: belief
8 Aug 2026 The Cognitive Revolution Thinking in Silico: Goodfire CTO Dan Balsam on Concept Manifolds & a $1000/Month ML Research Agent
“You know? And one example of that in my mind would be, like, don't train agents to maxim you know, with a reward signal that's, how much money they made on the Internet Yeah.”
Source trail
Everything needed to verify it.
- Speaker
- Nathan Labenz
- Attribution
- Verified speaker
- Claim type
- belief
- Recorded
- 8 Aug 2026
- Publisher
- The Cognitive Revolution
Transcript context
…Right. Fast forward to today, and it seems like hyperscaling RLVR is, not going super well. Right? We we sort of are seeing that we have problems arising from models just being, like, so tenacious in their pursuit of these goals. Do you have any sense of, like, you know, if you if we're gonna say, okay. Well, I talked to Zvi a couple days ago, and his basic take is, yeah. We're probably just gonna need compute limits. But I also feel like there might be some agreements around training techniques that we might all ought to say, you know, maybe not never, but, like, not now. You know? And one example of that in my mind would be, like, don't train agents to maxim you know, with a reward signal that's, how much money they made on the Internet Yeah. You know, in an in an open and, you know, potentially competitive or adversarial environment. Right? That's a recipe to get bad agents. I think there's a there's a big space of bad ideas. Yeah.…
Stored transcript either side of the excerpt. The highlighted words are the published quote; the surrounding text is unedited source, never generated.