High Signal Podcasts Evidence ledger
Method
Browse
← Back to evidence

Evidence receipt / prediction

Published · transcript-backed

Speaker unverified: prediction

13 Dec 2024 Latent Space Windsurf: The Enterprise AI IDE - with Varun and Anshul of Codeium AI

“works without, like, any, like, intent-based information, sure, that can become, like, fully agentic. And, like, we will learn what those tasks are, like, pretty quickly because we have a lot of data.”

— Speaker unverified

Source trail

Everything needed to verify it.

Speaker
Speaker unverified
Attribution
Not verified from this transcript
Claim type
prediction
Recorded
13 Dec 2024
Publisher
Latent Space

Transcript context

…I think it's like, how do you get there in a real principled manner? We obviously have Enterprise asking us all the time, like, oh, when's it going to be like end-to-end work? The reality is like, okay, well, if we have something in the ID that, again, can see your entire actions and get a lot of intent that you can't actually get if you're not in the ID, I mean, if the agent there has to always get human, like, involvement to keep on fixing itself, it's probably not ready to become a full end-to-end automated system because then we're just going to turn into a linter where, like, it produces a bunch of things and no one looks at any of it. Like, that's not the great end state. But if we start seeing, like, oh, yeah, there's common patterns that people do that, like, never require human involvement, just end-to-end just totally works without, like, any, like, intent-based information, sure, that can become, like, fully agentic. And, like, we will learn what those tasks are, like, pretty quickly because we have a lot of data. works without, like, any, like, intent-based information, sure, that can become, like, fully agentic. And, like, we will learn what those tasks are, like, pretty quickly because we have a lot of data. Maybe add on to that, I think that if the answer is, like, full agentic is called, like, is Devin, I think, like, yes, the answer is this product should become fully agentic and limited human interaction is the goal, is 100% the goal. And I think, honestly, of all usable products right now, I think we're the closest right now, of all usable products in an ID. Now, let me caveat this by saying I think there are lots of hard problems that have yet to be solved that we need to go out and solve to actually make this happen. Like, for instance, I think one of the most annoying parts about the product is the fact that you need to accept every command that kind of gets run. It's actually fairly annoying. I would like it to go out and run it. Unfortunately, me going out and running arbitrary binaries has some problems in that if it, like, RMRs my hard disk, I'm not going to be... It's actually a virus. I'm not saying, actually, the hacker needs to be with you. Yeah, it does become a virus. I think this is solvable with, like, with complex systems. I think we love working on complex systems infrastructure. I think we'll solve it. Now, the simpler way to go about solving this is don't run it on the user's machine and run it somewhere else because then if you bought that machine, you're kind of totally fine. I think, though, maybe there's a little bit of trade-off of, like, running it locally versus remotely, and I think we might change our mind on this, but I think the goal for this is not for this to be the final state. I think the goal for this is, A, it's actually able to do very complex tasks with limited human interaction, but it needs to know when to actually go back to the human, right? Also, on top of that, compress every cycle that the agent is running. Right now, actually, I even feel like the product is too slow for me sometimes right now. Even with it running really fast, it's objectively pretty fast, I would still want it to be faster, right? So there is, like, systems work and probably modeling work that needs to happen there to make the product even faster on both the retrieval side and the generation side, right? And then finally speaking, I think another key piece here that's, like, really important is I actually think asking people to do things explicitly is probably going to be more of an anti-pattern if we can actually go and passively suggest the entire change for the user. So almost imagine, as the user is using the product, that we're going to suggest the remainder of the PR without the user kind of, like, even asking us for it. I think this is sort of the beginning of the process. I think this is sort of the beginning of the process. But, yeah, like, these are hard problems. I can't give a particular deadline for this. I think this is, like, a big step up than what we had particularly in the past. But I think what Anshul said is 100% true, but the goal is for us to get better at this. I mean, the remote execution thing is interesting. You've wrote a post about the end of local host. Yeah. It's almost like then we were kind of like, well, no, maybe we do need the internet and, like, people want to run things. But now it's like, okay, no, actually, I don't really care. Like, I want the model to do the thing. And if you were like, you can do a task end-to-end, but it needs to run remotely, not on your computer, I'm sure most people would say, yeah.…

Stored transcript either side of the excerpt. The highlighted words are the published quote; the surrounding text is unedited source, never generated.

Search evidence