I've been using OpenAI operator for some time - but more and more websites are blocking it, such as LinkedIn and Amazon. That's two key use-cases gone (applying to jobs and online shopping). Operator is pretty low-key, but once Agent starts getting popular, more sites will block it. They'll need to allow a proxy configuration or something like that.
If people will actually pay for stuff (food, clothing, flights, whatever) through this agent or operator, I see no reason Amazon etc would continue to block them.
ChatGPT agent: bridging research and action
371–380 of 508 posts
Re: ChatGPT agent: bridging research and action
#372Earlier quoted context omitted.
I think the question then is what's the human error rate... We know we're not perfect... So if you're 100% rested and only have to find the edge case bug, maybe you'll usually find it vs you're burned out getting it 98% of the way there and fail to see the 2% of the time bugs... Wording here is tricky to explain but I think what we'll find is this helps us get that much closer... Of course when you spend your time bu…
Would be insane to expect an ai to just match us right…nooooo if it pertains computers/automation/ai it needs to be beyond perfect.
Re: ChatGPT agent: bridging research and action
#373Earlier quoted context omitted.
> how it normally takes him 4 to 8 hours to put together complicated, data-heavy reports. Now he fires off an agent request, goes to walk his dog, and comes back to a downloadable spreadsheet of dense data, which he pulls up and says "I think it got 98% of the information correct... This is where the AI hype bites people. A great use of AI in this situation would be to automate the collection and checking of data. Se…
“The quip about 98% correct should be a red flag for anyone familiar with spreadsheets” I disagree. Receiving a spreadsheet from a junior means I need to check it. If this gives me infinite additional juniors I’m good. It’s this popular pattern of HN comments - expect AI to behave deterministically correct - while the whole world operates on stochastically correct all the time…
You went from “do it again” to “go check the newbies work”.
To get to that stage your degree of proficiency would be “can make out which font is wrong at a glance.”
You wouldn’t be looking at the sheet, you would be running the model in your head.
That stopped being a stochastic function, with the error rate dropping significantly - to the point that making a mistake had consequences tacked on to it.
Re: ChatGPT agent: bridging research and action
#374For me the most interesting example on this page is the sticker gif halfway down the page. Up until now, chatbots haven't really affected the real world for me†. This feels like one of the first moments where LLMs will start affecting the physical world. I type a prompt and something shows up at my doorstep. I wonder how much of the world economy will be driven by LLM-based orders in the next 10 years. † yes I'm awar…
It went viral more than a year ago, so maybe you've seen it. On the Ritual Industries instagram, Brian (the guy behind RI) posted a video where he gives voice instruction to his phone assistant, which put the text through chatgpt, which generated openscad code, which was fed to his bambu 3d printer, which successfully printed the object. Voice to Stuff. I don't have ig anymore so I can't post the link, but it's easy…
Re: ChatGPT agent: bridging research and action
#375Earlier quoted context omitted.
In my experience the value of junior contributors is that they will one day become senior contributors. Their work as juniors tends to require so much oversight and coaching from seniors that they are a net negative on forward progress in the short term, but the payoff is huge in the long term.
I don't see how this can be true when no one stays at a single job long enough for this to play out. You would simply be training junior employees to become senior employees for someone else.
Re: ChatGPT agent: bridging research and action
#376Earlier quoted context omitted.
Given how well it seems to be going in those specific areas, it seems like it's more of a regulatory issue than a technological one.
This is a big moving of the goalposts. The optimists were saying Level 5 would be purchasable everywhere by ~2018. They aren’t purchasable today, just hail-able. And there’s a lot of remote human intervention. And San Francisco doesn’t get snow.
Or cows sharing the thoroughfares.
It should be obvious to all HNers that have lived or travelled to developing / global south regions - driving data is cultural data.
You may as well say that self driving will only happen in countries where the local norms and driving culture is suitable to the task.
A desperately anemic proposition compared to the science fiction ambition.
I’m quietly hoping I’m going to be proven wrong, but we’re better off building trains, than investing in level 5. It’s going to take a coordination architecture owned by a central government to overcome human behavior variance, and make full self driving a reality.
Re: ChatGPT agent: bridging research and action
#377Very slightly impressed by their emphasis on the gigantic (my word, not theirs) risk of giving the thing access to real creds and sensitive info.
But since people can cancel transactions with a credit card, that's what people are going to do, and it will be a huge mess every time.
Re: ChatGPT agent: bridging research and action
#378Earlier quoted context omitted.
Given how well it seems to be going in those specific areas, it seems like it's more of a regulatory issue than a technological one.
Ah, those pesky regulations that try to prevent road accidents... If it's not a technological limitation, why aren't we seeing self-driving cars in countries with lax regulations? Mexico, Brazil, India, etc. Tesla launched FSD in Mexico earlier this year, but you would think companies would be jumping at the opportunity to launch in markets with less regulation. So this is largely a technological limitation. They hav…
Re: ChatGPT agent: bridging research and action
#379The "spreadsheet" example video is kind of funny: guy talks about how it normally takes him 4 to 8 hours to put together complicated, data-heavy reports. Now he fires off an agent request, goes to walk his dog, and comes back to a downloadable spreadsheet of dense data, which he pulls up and says "I think it got 98% of the information correct... I just needed to copy / paste a few things. If it can do 90 - 95% of the…
Re: ChatGPT agent: bridging research and action
#380Earlier quoted context omitted.
Given how well it seems to be going in those specific areas, it seems like it's more of a regulatory issue than a technological one.
Maybe, but it's also going to be a financial issue eventually too My city had Car2Go for a couple of years, but it's gone now. They had to pull out of the region because it wasn't making them enough money I expect Waymo and any other sort of vehicle ridesharing thing will have the same problem in many places