Evidence receipt / evaluation
Published · transcript-backedSuhail Doshi: evaluation
20 Dec 2023 Latent Space The AI-First Graphics Editor - with Suhail Doshi of Playground AI
“We're still at the beginning of building like the best benchmark we can that aligns most with just user happiness, I think, because we're not we're not like putting these in papers and trying to like win, you know, I don't know, awards at ICCV or something if they have awards.”
Source trail
Everything needed to verify it.
- Speaker
- Suhail Doshi
- Attribution
- Verified speaker
- Claim type
- evaluation
- Recorded
- 20 Dec 2023
- Publisher
- Latent Space
Transcript context
…They should embed to the same space. Yeah. And just like all these very interesting weirdo things. And so we have so many of these and then we kind of like evaluate whether the models are any good at it. And the reality is that they're all bad at it. And so then you're just picking the most aesthetic image. We're still at the beginning of building like the best benchmark we can that aligns most with just user happiness, I think, because we're not we're not like putting these in papers and trying to like win, you know, I don't know, awards at ICCV or something if they have awards. You could. That's absolutely a valid strategy.…
Stored transcript either side of the excerpt. The highlighted words are the published quote; the surrounding text is unedited source, never generated.