High Signal Podcasts Evidence ledger
Method
Browse
← Back to evidence

Evidence receipt / belief

Published · transcript-backed

Speaker unverified: belief

5 Aug 2025 Machine Learning Street Talk DeepMind Genie 3 [World Exclusive] (Jack Parker Holder, Shlomi Fruchter)

“I think that's what you're alluding to. So I think in in the case of we're still in a similar place, I would say, when we think about the world models because you have to provide the description of the world that you want to to maybe walk in into an experience.”

— Speaker unverified

Source trail

Everything needed to verify it.

Speaker
Speaker unverified
Attribution
Not verified from this transcript
Claim type
belief
Recorded
5 Aug 2025
Publisher
Machine Learning Street Talk

Transcript context

…Yeah. I think I think this is the fundamental thing because certainly with creative models at the moment, like weirdly, counterintuitively, you need more skill to make it do something interesting than you did before. Like the average creative process now for someone designing a thumbnail on YouTube is they they mix together, you know, they might use the contact model, they might use an upscale, they might then like, you know, use another image generator model. You get this huge kind of compositional tree of operations that happen. And it's very, very highly skilled because a lot of the kind of structure for constraining the generation of these models still comes from our own abstract understanding of the world. And this is kind of what Kenneth Stanley was saying in a sense. He was saying that we have this understanding of the world, which is constrained by things like symmetry and, you know, various different rules. And we then kind of we we hint to the model. We we constrain the model in the prompt using those things. Would the models ever be able to do that without the humans needing to prompt them? So I I think what's interesting is that eventually, what we find like, what humans find interesting and worth, you know, maybe watching or investigating or researching, it's it's eventually being defined by people. And and I think in the case, for example, if it's video, if it's it's video generation, for example, then we see that people go and and find ways that maybe we weren't like, they use the tool that we put in front of them to generate new things. So for example, we have people try to cut like, we we have the ASMR videos of people cutting, you know, fruits made of glass. Right? Which is not something you can do in the real world, and the novelty comes from from the prompt, basically. I think that's what you're alluding to. So I think in in the case of we're still in a similar place, I would say, when we think about the world models because you have to provide the description of the world that you want to to maybe walk in into an experience. But some elements are not like, would kinda, like, emerge from and will be inferred from the prompt that you provide. Right? So you can maybe write a very short prompt, but still the world will have much more richness. So I think there's a question of where does this rich richness coming from, and I think where it's different maybe levels of of the the the ability of models to to bring this rich richness into your experience. But I think it's kinda like over time, we see that it becomes more and more higher and higher, and little information provided by users can actually generate very rich videos or experiences. So I would say it's not like a it's a bit of a evolving answer, I would say. Like, over time, I expect that more inputs to the model or you can think about it like the the the person is providing a seed, and from that seed, we can maybe generate more, like, more elaborate descriptions and finally an experience. So I don't think about it as, like, a 1 step process, but more of, like, a series of of creative steps. Or each 1 of them can be it can happen by can be done by a person or by an AI model, and together, they generate maybe something new. Yeah. And and that's what we're seeing play out on Twitter. That's you know, because the creative process is like, you know, generate, discriminate, generate, discriminate. And we memetically share all of the prompts that work. And that's why we've just created this beautiful phylogeny of creative artifacts that are exploring the, you know, the the space of of these models, which is beautiful. And I'm thinking about the future. I mean, I know you probably can't speculate about this, but this could be the next YouTube. It could be a new form of virtual reality. You know, in philosophy, there's this thing called the experience machine, where you you plug yourself into this better than life matrix simulation. And no 1 wants to leave the experience machine because it's better than real life. But we could co create something like that. Right? We could we could have it could be on a on a phone or a virtual headset. And we could create these worlds and portals between the worlds, and it would just be a never ending simulation?…

Stored transcript either side of the excerpt. The highlighted words are the published quote; the surrounding text is unedited source, never generated.

Search evidence