Earlier quoted context omitted.
What tasks are you running that take more than a few minutes without intervention?
When using spec writter and sub-tasking tools like TaskMaster, Kiro, etc. I've experienced Claude Code to take 30-60+ minutes for a more complex feature
ChatGPT agent: bridging research and action
411–420 of 508 posts
Re: ChatGPT agent: bridging research and action
#412Re: ChatGPT agent: bridging research and action
#413Re: ChatGPT agent: bridging research and action
#414Earlier quoted context omitted.
In the system I'm building the main agent doesn't have access to tools and must call scoped down subagents who have one or two tools at most and always in the same category (so no mixed fetch and calendar tools). They must also return structured data to the main agent. I think that kind of isolation is necessary even though it's a bit more costly. However since the subagents have simple tasks I can use super cheap mo…
What isolation is there? If a compromised sub agent returns data that gets inserted into the main agents context (structured or not) then the end result is the same as if the main agent was directly interacting with the compromising resource is it not?
Re: ChatGPT agent: bridging research and action
#415The "spreadsheet" example video is kind of funny: guy talks about how it normally takes him 4 to 8 hours to put together complicated, data-heavy reports. Now he fires off an agent request, goes to walk his dog, and comes back to a downloadable spreadsheet of dense data, which he pulls up and says "I think it got 98% of the information correct... I just needed to copy / paste a few things. If it can do 90 - 95% of the…
And why 98%? Why not 99% right? Or 99.9% right? I know they can't outright say 100% because everyone knows that's a blatant lie, but we're okay with them bullshitting about the 98% number here?
Also there's no universe in which this guy gets to walk his dog while his little pet AI does his work for him, instead his boss is going to hound him into doing quadruple the work because he's now so "efficient" that he's finishing his spreadsheet in an hour instead of 8 or whatever. That, or he just gets fired and the underpaid (or maybe not even paid) intern shoots off the same prompt to the magic little AI and does the same shoddy work instead of him. The latter is definitely what the C-suite is aiming for with this tech anyway.
Re: ChatGPT agent: bridging research and action
#416Earlier quoted context omitted.
The proper use of these systems is to treat them like an intern or new grad hire. You can give them the work that none of the mid-tier or senior people want to do, thereby speeding up the team. But you will have to review their work thoroughly because there is a good chance they have no idea what they are actually doing. If you give them mission-critical work that demands accuracy or just let them have free rein with…
What a awful way to think about internship. The goal is to help people grow, so they can achieve things they would not have been able to deal with before gaining that additional experience. This might include boring dirty work, yes. But that means they thus prove they can overcome such a struggle, and so more experienced people should be expected to also be able to go though it - if there is no obvious more pleasant…
Re: ChatGPT agent: bridging research and action
#417Earlier quoted context omitted.
THIS is the main problem. I was listening the whole time for them to announce a way to run it locally or at least proxy through your local devices. Alas the Deepseek R1 distillation experience they went through (a bit like when Steve Jobs was fuming at Google for getting Android to market so quickly) made them wary of showing to many intermediate results, tricks etc. Even in the very beginning Operator v1 was unable…
This is why an on device browser is coming. It'll let the AI platforms get around any other platform blocks by hijacking the consumer's browser. And it makes total sense, but hopefully everyone else has done the game theory at least a step or two beyond that.
Re: ChatGPT agent: bridging research and action
#418Re: ChatGPT agent: bridging research and action
#419Is it available in the EU yet? Doesn't look like...
Re: ChatGPT agent: bridging research and action
#420Very slightly impressed by their emphasis on the gigantic (my word, not theirs) risk of giving the thing access to real creds and sensitive info.