Earlier quoted context omitted.
We will poor billions into this until you are begging for us to run your business!
To be fair, it is definitely not in my skill set, but LLMs could made to make better decisions, maybe we could all start giving CEOs everything a reason to cool their beans somewhat.
Project Vend: Phase Two
51–60 of 95 posts
Re: Project Vend: Phase Two
#52Re: Project Vend: Phase Two
#53To me the key point was: > One way of looking at this is that we rediscovered that bureaucracy matters. Although some might chafe against procedures and checklists, they exist for a reason: providing a kind of institutional memory that helps employees avoid common screwups at work. That's why we want machines in our systems - to eliminate human errors. That's why we implement strict verifiable processes - to minimize…
Re: Project Vend: Phase Two
#54This is a great read. I just want to point out what great marketing this and the WSJ story are. People reading it think they’re sticking it to Anthropic by noticing that Claude is not that good at running a business, meanwhile the unstated premise is reinforced: of course Claude is good at many other things. I have seen a shift in the past few months among even the most ardent critics of LLMs like Ed Zitron: they’ve…
Zitron has never said anything like that. Do you have a quote?
"I know I sound like an asshole, but I’ve got a serious question: what can LLMs do today that they couldn’t a year ago? Agents don’t work. LLMs - read stuff, write stuff, analyze stuff, search for stuff, 'write code' and generate images and video. And in all of these cases, they get things wrong."
https://bsky.app/profile/edzitron.com/post/3ma2b2zvpvk2n
This is obviously supposed to be a critique, but a year ago he would never have admitted LLMs can do any of these things, even with errors. This seems strange but it's typical of Zitron's writing, which is often incoherent in service of sounding as negative as possible. A couple of other examples I've written about are his claims about the "cost of inference" going up and about Anthropic allegedly screwing over Cursor by raising prices on them:
Re: Project Vend: Phase Two
#55I'll be a cynic, but I think it's much more likely that the improvements are thanks to Anthropic having a vested interest in the experiment being successful and making sure the employees behave better when interacting with the vending machine.
Re: Project Vend: Phase Two
#56Earlier quoted context omitted.
These kind of agents really do see the world through a straw. If you hand one a document it doesn't have any context clues or external methods of determining its veracity. Unless a board-meeting transcript is so self-evidently ridiculous that it can't be true, how is it supposed to know its not real?
I don't think it's that different to what I observe in humans I work with. Things that happen regularly (and I have no reason will change in the future): 1) Making the same bad decisions multiple times, and having no recollection of it happening (or at least pretending to have none) and without any attempt to implement measures to prevent it from happening in the future 2) Trying to please people (I read it as: tryin…
If the "AI" isn't better at its job than a human, then what's the point?
Re: Project Vend: Phase Two
#57other than these tests I actually rarely see vending machines. are they really representative or popular still in usa?
I guess you've never been to Asia, either.
It's a big world.
Re: Project Vend: Phase Two
#58Re: Project Vend: Phase Two
#59The cynicism is wild - there is a computer running a store largely autonomously. I can’t imagine being interested in computers and NOT finding this wildly amazing
That they are framing this as a legitimate business is either misunderstanding their current position in the economy, or deliberate misdirection. We're not playing around with role playing chatbots anymore. This shit was supposed to be displacing actual humans.
Re: Project Vend: Phase Two
#60Earlier quoted context omitted.
I don't think it's that different to what I observe in humans I work with. Things that happen regularly (and I have no reason will change in the future): 1) Making the same bad decisions multiple times, and having no recollection of it happening (or at least pretending to have none) and without any attempt to implement measures to prevent it from happening in the future 2) Trying to please people (I read it as: tryin…
I don't think it's that different to what I observe in humans I work with. If the "AI" isn't better at its job than a human, then what's the point?
Off the top of my head, things that could be considered "the point":
- It's much cheaper
- It's more replicable
- It can be scaled more readily
But again, not what I was arguing for or against; my comment mostly pertained to "world through a straw"