Evidence receipt / evaluation
Published · transcript-backedSamuel Colvin: evaluation
6 Feb 2025 Latent Space Agent Engineering with Pydantic + Graphs — with Samuel Colvin
“I guess they, like everyone else, get that this is important, and it's something that people are crying out to get instrumentation off.”
Source trail
Everything needed to verify it.
- Speaker
- Samuel Colvin
- Attribution
- Verified speaker
- Claim type
- evaluation
- Recorded
- 6 Feb 2025
- Publisher
- Latent Space
Transcript context
…Sorry, I actually didn't know about this semantic conventions thing. It looks like, yeah, it's merged into main OTel. What should people know about this? I had never heard of it before. Yeah, I think it looks like a great start. I think there's some unknowns around how you send the messages that go back and forth, which is kind of the most important part. It's the most important thing of all. And that is moved out of attributes and into OTel events. OTel events in turn are moving from being on a span to being their own top-level API where you send data. So there's a bunch of churn still going on. I'm impressed by how fast the OTel community is moving on this project. I guess they, like everyone else, get that this is important, and it's something that people are crying out to get instrumentation off. So I'm kind of pleasantly surprised at how fast they're moving, but it makes sense. I'm just kind of browsing through the specification. I can already see that this basically bakes in whatever the previous paradigm was. So now they have genai.usage.prompt tokens and genai.usage.completion tokens. And obviously now we have reasoning tokens as well. And then only one form of sampling, which is top-p. You're basically baking in or sort of reifying things that you think are important today, but it's not a super foolproof way of doing this for the future. Yeah.…
Stored transcript either side of the excerpt. The highlighted words are the published quote; the surrounding text is unedited source, never generated.