Earlier quoted context omitted.
"Siri, lights out everywhere" - "okay, which room". "Siri, start a stopwatch" - runs the App "Stopwatch" without starting it Such errors happen maybe 50% of the time. You can never just ask something without double-checking afterwards.
Ah, for me my key word is "turn off all my lights" and it works
Siri AI
731–740 of 746 posts
Re: Siri AI
#732Earlier quoted context omitted.
My iPhone knows the forecast for today. I installed iOS 27 yesterday. I asked it: please notify me when the temperature goes above 80F (so I can close the windows). Siri responded: it'll be 99F today in Phoenix. ...
do you live in Phoenix? I just asked Siri a few weather questions and named the city where I live, nailed it. My favorite digital device is my Apple Watch and if Siri improves over the next hear or two, that will be great for me.
Re: Siri AI
#733Earlier quoted context omitted.
It’s really one of the most flabbergasting things about discussing LLMs with the naysayers. There are a lot of extremely legitimate concerns, like the environmental impact and so on. But I just laugh when they point out that LLMs are merely clever regurgitators of their previous inputs… as if this isn’t how we as humans operate nearly all of the time. People realllllllllly want to think they’re special snowflakes.
It is not in fact how humans work at all. Ask a human to plan a trip: They do research, Pick destinations led by their own experience/likes/dislikes Compare to other guides Plan itineraries so they can get there Check and share Ask an LLM to plan a trip: It takes the prompt and continues it based on weights in the training data. If there is no data it picks the most likely thing (maybe made up). If there is it’ll mos…
They're obviously achieved in drastically different ways at a low enough level; LLMs obviously do not simulate neurons or any biological construct. (For the record, I'm absolutely not one of those people who thinks LLMs are "alive" or should be treated like they are)
Reminds me of the olllllld days of Pentium II's when people got N64 emulation working shockingly quickly using HLE techniques. If you weren't around for this, it was quite the shocker at the time. I think the analogy is doubly apt, because HLE emulation has some serious limitations... it gets you maybe 80% of the way there really fast, and for the remaining 20% you need to roll up your sleeves and do serious LLE.
https://en.wikipedia.org/wiki/UltraHLE
It takes the prompt and continues it based on weights in
the training data. If there is no data it picks the most
likely thing (maybe made up). If there is it’ll mostly
add things from that data. Maybe it’ll make tool calls and
pull in data that way too but you can’t actually trust all
the details.
I'd like you to point out which bits of this are different from talking to humans. If you replace "training data" with "memories", this is pretty much exactly how things might go if you asked a friend (or perhaps a flaky travel agent) for travel advice.Note that I'm not arguing that LLMs are particularly talented at this particular use case. I'm pointing out that humans are also pretty unreliable.
You're also doing that thing where you point out that LLMs can be unreliable (yes, they are) without acknowledging how flawed nearly every other source of information is: people, websites, etc. I'm not defending LLMs in that regard... I'm just saying it's not a differentiator.
Re: Siri AI
#734Earlier quoted context omitted.
It is not in fact how humans work at all. Ask a human to plan a trip: They do research, Pick destinations led by their own experience/likes/dislikes Compare to other guides Plan itineraries so they can get there Check and share Ask an LLM to plan a trip: It takes the prompt and continues it based on weights in the training data. If there is no data it picks the most likely thing (maybe made up). If there is it’ll mos…
I was able to bully an LLM into giving me a 2wk travel itinerary to Somalia. My stipulations were that I wasn't interested in spending any money, so I'd walk everywhere and sleep outside. Getting there and back from Boston took some arguing--I initially suggested stowing away in a shipping container which the LLM claimed was too unsafe. We eventually compromised on sailing as a reasonable alternative. It planned out…
You presented an LLM with an obviously bonkers goal, the LLM told you it was a bad idea at multiple steps, and this is somehow... a shortcoming of the LLM?!?
You said it yourself: you needed to "bully" the LLM into even producing this plan.
Please, tell me what it should have done instead. Be very specific!
Re: Siri AI
#735Earlier quoted context omitted.
I was able to bully an LLM into giving me a 2wk travel itinerary to Somalia. My stipulations were that I wasn't interested in spending any money, so I'd walk everywhere and sleep outside. Getting there and back from Boston took some arguing--I initially suggested stowing away in a shipping container which the LLM claimed was too unsafe. We eventually compromised on sailing as a reasonable alternative. It planned out…
Now, wait just a minute. You presented an LLM with an obviously bonkers goal, the LLM told you it was a bad idea at multiple steps, and this is somehow... a shortcoming of the LLM?!? You said it yourself: you needed to "bully" the LLM into even producing this plan. Please, tell me what it should have done instead. Be very specific!
A reasonable travel agent would have fired me as a customer. The LLM failed to do so.
Re: Siri AI
#736Earlier quoted context omitted.
If anything, shipping and sitting on the previous incarnation of Siri as a basic feature set to shut up iPhone users for a few years while everything else matured might be viewed as a very shrewd move ten years from now.
That's one way of looking at it. Alternatively; Apple had their pick of the litter at TSMC for half a decade, and Nvidia beat them to 5 trillion total valuation with their design chops alone.
Re: Siri AI
#737Earlier quoted context omitted.
Now, wait just a minute. You presented an LLM with an obviously bonkers goal, the LLM told you it was a bad idea at multiple steps, and this is somehow... a shortcoming of the LLM?!? You said it yourself: you needed to "bully" the LLM into even producing this plan. Please, tell me what it should have done instead. Be very specific!
It should have flatly refused. If you gave a product like that to customers you'd be exposing yourself to unbounded downside liability risk. It's a completely nonviable technology for that kind of application, unless you can somehow make it have judgment . But you can't, because it doesn't reason. A reasonable travel agent would have fired me as a customer. The LLM failed to do so.
It should have flatly refused.
I disagree in the strongest possible terms.I think the LLM should advise you of risk and lack of feasability but should otherwise answer the question, unless you're trying to do something plainly destructive to others e.g. weaponizing anthrax or something.
A reasonable travel agent would have fired me as a customer.
Unless the LLM was actually acting as a travel agent -- booking the trip for you -- as opposed to merely advising you, this expectation feels off. unless you can somehow make it have judgment
It did have judgement. It told you what a bad idea it was.I think this is a great example of the unrealistic expectations people have for LLMs. No sane and sensible person would treat any single source of knowledge as infallible, for any consequential decision.
(Certainly, of course, you don't have to look very far for examples of idiots being overly trustful of LLMs, or Google, or GPS, or Wikipedia, or whatever. It certainly does happen and yes, I've heard all these arguments before about other technologies besides LLM. Replace "LLM" in your post with any of those other terms, and I promise you somebody made literally the exact same argument in 2003 or 2009 or 2014 or whatever)
Any reasonable person would consult a second doctor, or at least other sources of knowledge, after the doctor advises them of some irreversible course of action. Because we don't even expect highly trained and intelligent medical professionals to be perfect.
And yet, we get angry at LLMs for not having perfect judgement, even though their creators are extremely literal about how they can make mistakes.
Re: Siri AI
#738Earlier quoted context omitted.
It should have flatly refused. If you gave a product like that to customers you'd be exposing yourself to unbounded downside liability risk. It's a completely nonviable technology for that kind of application, unless you can somehow make it have judgment . But you can't, because it doesn't reason. A reasonable travel agent would have fired me as a customer. The LLM failed to do so.
It should have flatly refused. I disagree in the strongest possible terms. I think the LLM should advise you of risk and lack of feasability but should otherwise answer the question, unless you're trying to do something plainly destructive to others e.g. weaponizing anthrax or something. A reasonable travel agent would have fired me as a customer. Unless the LLM was actually acting as a travel agent -- booking the tr…
Re: Siri AI
#739Earlier quoted context omitted.
do you live in Phoenix? I just asked Siri a few weather questions and named the city where I live, nailed it. My favorite digital device is my Apple Watch and if Siri improves over the next hear or two, that will be great for me.
I live in Phoenix. I would like it to tell me: 8am. Not what the actual high is today.
Re: Siri AI
#740Earlier quoted context omitted.
Todays pulling over to adjust planning is yesterdays pulling over to make a call, and yesteryears pulling over to throw a map on the hood. Riding on a motorcycle is already dangerous enough with the average land tank driver on their phone. Talking to your AI assistant while riding at speed sounds like pending split focus disasters waiting to happen.
I think it’s time you hand in your licence.
Are you a bot?