Live data from Hacker News

We are beginning to roll out new voice and image capabilities in ChatGPT

openai.com

201–210 of 914 posts

Re: We are beginning to roll out new voice and image capabilities in ChatGPT

#201

I like how they silently removed the web browsing (Bing browsing) chat feature after first having it disabled for several months. A proper notice about them removing the feature would've been nice. Maybe I missed it (someone please correct me if wrong), but the last I heard officially it was temporarily disabled while they fix something. Next thing I know, it's completely gone from the platform without another peep.

Yes, that was a disappointment, and I agree it looks like they aren't going to re-enable it anytime soon. However I find that Perplexity AI does a better job of using web search than ChatGPT ever did, and I use it more than ChatGPT for that reason.

Re: We are beginning to roll out new voice and image capabilities in ChatGPT

#202
post #49

Earlier quoted context omitted.

I also don't believe LLMs are "conscious", but I also don't know what that means, and I have yet to see a definition of "statistically guessing next word" that cannot be applied to what a human brain does to generate the next word.

I believe that the distinguishing factor between what an LLM and a human brain do to generate the next word is that the human brain expresses intentionality originating from inner states and future expectations. As I type this comment I'm sure one could argue that the biological neural networks in my brain are choosing the next word based on statistical guessing, and that the initial prompt was your initial comment.…

>> the human brain expresses intentionality originating from inner states and future expectations

How is this different from and/or the same as the concept of "attention" as used in transformers?

Re: We are beginning to roll out new voice and image capabilities in ChatGPT

#203

My biggest complaint with OpenAI/ChatGPT is their horrible "marketing" (for lack of a better term). They announce stuff like this (or like plugins), I get excited, I go to use it, it hasn't rolled out to me yet (which is frustrating as a paying customer), and my only recourse is.... check back daily? They never send an email "Plugins are available for you!", "Voice chat is now enabled on your account!" and so often I…

They do explain why in the post. (Still, you may not agree, of course.)

> We are deploying image and voice capabilities gradually > > OpenAI’s goal is to build AGI that is safe and beneficial. We believe in making our tools available gradually, which allows us to make improvements and refine risk mitigations over time while also preparing everyone for more powerful systems in the future. This strategy becomes even more important with advanced models involving voice and vision.

Re: We are beginning to roll out new voice and image capabilities in ChatGPT

#204
post #49

Earlier quoted context omitted.

I also don't believe LLMs are "conscious", but I also don't know what that means, and I have yet to see a definition of "statistically guessing next word" that cannot be applied to what a human brain does to generate the next word.

I believe that the distinguishing factor between what an LLM and a human brain do to generate the next word is that the human brain expresses intentionality originating from inner states and future expectations. As I type this comment I'm sure one could argue that the biological neural networks in my brain are choosing the next word based on statistical guessing, and that the initial prompt was your initial comment.…

I think one mean difference in LLM, is what Micheal Scott said in The Office: "Sometimes I'll start a sentence, and I don't even know where it's going. I just hope I find it along the way. Like an improv conversation. An improversation"

Human will know what they want to express, choosing words to express it might be similar to LLM process of choosing words, but for LLM it doesn't have that "Here is what i know to express part", i guess that the conscious part?

Re: We are beginning to roll out new voice and image capabilities in ChatGPT

#205
post #129

The most important question for me: did it stop inventing facts?

> In particular, beta testers expressed concern that the model can make basic errors, sometimes with misleading matter-of-fact confidence. One beta tester remarked: “It very confidently told me there was an item on a menu that was in fact not there.” However, Be My Eyes was encouraged by the fact that we noticeably reduced the frequency and severity of hallucinations and errors over the time of the beta test. In particular, testers noticed that we improved optical character recognition and the quality and depth of descriptions.

So no, but maybe less than it used to?

Re: We are beginning to roll out new voice and image capabilities in ChatGPT

#206

Earlier quoted context omitted.

This is what effectively doctors do - educated guessing. In my view, while statistical models would probably be an improvement ( assuming all confounding factors are measured ), the ultimate solution is not to get better at educated guessing, but to remove the guessing completely, with diagnostic tests that measure the relevant bio-medical markers.

Good tests This becomes even more true when you consider there is risk to every test. Some tests have obvious risks (radiation risk from CT scans, chance of damage from spinal fluid tap). Other tests the risk is less obvious (sending you for a blood test and awaiting the results might not be a good idea if that delays treatment for some ailment already pretty certain). In the bigger picture, any test that costs money…

Sure tests cost money - and today there is a funnel pathway - the educated guess is a funnel/filter where the next step which is often a biomedical test/investigation.

But if we are talking about being truly transformative - then a Star-trek tricorder is the ultimate goal, rather than a better version of twenty questions in my view.

So I'm not saying it's not useful, just that it's not the ultimate solution.

Re: We are beginning to roll out new voice and image capabilities in ChatGPT

#207
post #56

Earlier quoted context omitted.

> The only downside is that it now make me feel bad that I'm not doing anything with it yet. If that's the only downside that you see... I guess enhanced phishing/impersonation and all the blackhat stuff that come with it don't count. I for one already miss the time where companies had support teams made of actual people.

I work as a ethical hacker, so I'm well aware of the phishing and impersonation possibilities. But the net positive is so, so much bigger for society that I'm sure we'll figure it out. And yes, in 20 years you can tell your kids that 'back in my day' support consisted of real people. But truthfully, as someone who worked on a ISP helpdesk it's much better for society if these people move on to more productive areas.

> But the net positive is so, so much bigger for society that I'm sure we'll figure it out.

Considering that the democratic backsliding across the globe is coincidentally happening at the same time as the rise of social media and echo chambers, are we sure about that? LLM have the opportunity to create a handcrafted echo chamber for every person on this planet, which is quite risky in an environment where almost every democracy of the planet is fighting against radical forces trying to abolish it.

Re: We are beginning to roll out new voice and image capabilities in ChatGPT

#208
post #49

Earlier quoted context omitted.

I also don't believe LLMs are "conscious", but I also don't know what that means, and I have yet to see a definition of "statistically guessing next word" that cannot be applied to what a human brain does to generate the next word.

I believe that the distinguishing factor between what an LLM and a human brain do to generate the next word is that the human brain expresses intentionality originating from inner states and future expectations. As I type this comment I'm sure one could argue that the biological neural networks in my brain are choosing the next word based on statistical guessing, and that the initial prompt was your initial comment.…

You make a good point. I would not equate consciousness to intentionality though.

One of the big problems with discussions about AI and AI dangers in my mind is that most people conflate all of the various characteristics and capabilities that animals like humans have into one thing. So it is common to use "conscious", "self-aware", "intentional", etc. etc. as if they were all literally the same thing.

We really need to be able to more precise when thinking about this stuff.

Re: We are beginning to roll out new voice and image capabilities in ChatGPT

#209

My biggest complaint with OpenAI/ChatGPT is their horrible "marketing" (for lack of a better term). They announce stuff like this (or like plugins), I get excited, I go to use it, it hasn't rolled out to me yet (which is frustrating as a paying customer), and my only recourse is.... check back daily? They never send an email "Plugins are available for you!", "Voice chat is now enabled on your account!" and so often I…

We're heading for the singularity and you're complaining about marketing?

Re: We are beginning to roll out new voice and image capabilities in ChatGPT

#210
post #49

Earlier quoted context omitted.

I also don't believe LLMs are "conscious", but I also don't know what that means, and I have yet to see a definition of "statistically guessing next word" that cannot be applied to what a human brain does to generate the next word.

I keep feeling that consciousness is a bit of a red herring when it comes to AI. People have intuitions that things other than humans cannot develop consciousness which they then extrapolate to thinking AI can't get past a certain intelligence level. In fact my view is that consciousness is just a mysterious side effect of the human brain, and is completely irrelevant to the behaviour of a human. You can be intellige…

Unless you think that consciousness is entirely a post hoc process to rationalize thoughts already had and decisions already made, which is very much unlike how most people would describe their experience of it, I don't see how you could possibly say that it is irrelevant to the behavior of a human.
Post reply on HN