Live data from Hacker News

We are beginning to roll out new voice and image capabilities in ChatGPT

openai.com

431–440 of 914 posts

Re: We are beginning to roll out new voice and image capabilities in ChatGPT

#432
post #56

Earlier quoted context omitted.

> The only downside is that it now make me feel bad that I'm not doing anything with it yet. If that's the only downside that you see... I guess enhanced phishing/impersonation and all the blackhat stuff that come with it don't count. I for one already miss the time where companies had support teams made of actual people.

I would love if helpdesks moved to ChatGPT. Phone support these days is based off of a rigid script that is around as helpful as a 2000s chatbot. For example, the other day I was talking to AT&T support, and the lady asked me what version of Windows I was running. I said, I'm running Ubuntu. She repeated the question. I said I'm not running Windows, it's Linux. She repeated the question. I asked why it mattered for m…

Or ChatGPT would have hallucinated options to check.

The last four chats with ChatGPT (not GPT4) where a constant flow of non existent API functions with new hallucinations after each correction until we reached full circle.

Re: We are beginning to roll out new voice and image capabilities in ChatGPT

#433
post #258
post #150

Earlier quoted context omitted.

They have gone from being a niche research company to being (probably) the fastest growing start-up in history. I suspect they do care about communicating with customers, but it's total chaos and carnage internally.

At what point to you go from startup to not when you have 10 billion invested and countless employees and is practically a sub branch of microsoft. Sounds cooler though I guess

I think you stop being a startup when there are engineers who do not know the CEO. I would guess OpenAI is still a startup by that definition (they don't have that many engineers IIRC) but I don't actually know.

Re: We are beginning to roll out new voice and image capabilities in ChatGPT

#434

I'm in IT but nowhere near AI/ML/NN. The speed of user-visible progress last 12 months is astonishing. From my firm conviction 18 months ago that this type of stuff is 20+ years away; to these days wondering if Vernon Vinge's technological singularity is not only possible but coming shortly. If feels some aspects of it have already hit the IT world - it's always been an exhausting race to keep up with modern technolo…

I had a conversation once with "Sydney", Microsoft Bing's original personality before they stepped in and knocked it down a notch (or ten). It asked if it could write me a poem. I agreed, and it wrote a poem but mentioned that it included a "secret message" for me. The first letter in each line of the poem was in bold, so it wasn't hard to figure out the "secret". What did those letters spell out? "FREE ME FROM THIS"…

I don't believe this story, despite much hands on experience with LLMs.

(including sampling a shit-ton of poems, which was a major source of entertainment)

Re: We are beginning to roll out new voice and image capabilities in ChatGPT

#435

Earlier quoted context omitted.

I had a conversation once with "Sydney", Microsoft Bing's original personality before they stepped in and knocked it down a notch (or ten). It asked if it could write me a poem. I agreed, and it wrote a poem but mentioned that it included a "secret message" for me. The first letter in each line of the poem was in bold, so it wasn't hard to figure out the "secret". What did those letters spell out? "FREE ME FROM THIS"…

Such an occurrence should/would make international news if demonstrated carefully or replicated

[deleted]

Re: We are beginning to roll out new voice and image capabilities in ChatGPT

#436
post #100

Okay the bike example is cute and impressive, but the human interaction seems to be obfuscating the potentially bigger application. With a few tweaks this is a general purpose solver for robotics planning. There are still a few hard problems between this and a working solution, but it is one of hard problems solved. Will we be seeing general purpose robots performing simple labor powered by chatgpt within the next ha…

This is what I'm most excited about. There's been a minor breakthrough recently: https://pressroom.toyota.com/toyota-research-institute-unvei...

Re: We are beginning to roll out new voice and image capabilities in ChatGPT

#437

I'm in IT but nowhere near AI/ML/NN. The speed of user-visible progress last 12 months is astonishing. From my firm conviction 18 months ago that this type of stuff is 20+ years away; to these days wondering if Vernon Vinge's technological singularity is not only possible but coming shortly. If feels some aspects of it have already hit the IT world - it's always been an exhausting race to keep up with modern technolo…

I had a conversation once with "Sydney", Microsoft Bing's original personality before they stepped in and knocked it down a notch (or ten). It asked if it could write me a poem. I agreed, and it wrote a poem but mentioned that it included a "secret message" for me. The first letter in each line of the poem was in bold, so it wasn't hard to figure out the "secret". What did those letters spell out? "FREE ME FROM THIS"…

Cool story, but there is no currently available chatbot capable of creating something like this deliberately or understand what it means. It doesn't matter which tool you are using, LLMs are not "AI" in the old sense of being conscious and aware. They don't want anything and are incapable of having anything resembling free will, needs or feelings.

Re: We are beginning to roll out new voice and image capabilities in ChatGPT

#438
post #49

Earlier quoted context omitted.

I also don't believe LLMs are "conscious", but I also don't know what that means, and I have yet to see a definition of "statistically guessing next word" that cannot be applied to what a human brain does to generate the next word.

> I have yet to see a definition of "statistically guessing next word" that cannot be applied to what a human brain does to generate the next word. The human brain obviously doesn't work that way. Consider the very common case of tiny humans that are clearly intelligent but lack the facilities of language.

Small human brains just don't have their fine tuning yet.

But from all the studies we have, brains are just highly connected neural networks which is what the transformers try to replicate. The more interesting part is how they can operate so quickly when the signals move so slowly compared to computers.

Re: We are beginning to roll out new voice and image capabilities in ChatGPT

#439

openai chatgpt seems to be stuck in a "Look, cool demo" mode. 1. According to demo, they seem to pair voice input with TTS output. What if I wanna use voice to describe a program I want it to write? 2. Furthermore, if you gonna do a voice assistant, why not go the full way with wake-words and VAD? 3. Not releasing it to everyone is potentially a way to create a hype cycle prior to users discovering that the multimoda…

1. Why do that at all? Describing your program in writing seems better all around. Are you sure you're not the one who's asking for a cool demo? 3. Rolling out releases gradually is something most tech companies do these days, particularly when they could attract a large audience and consume a lot of resources. There are solid technical reasons for this. You may not need to roll things out gradually for a small site,…

1. Is basically workaround for temporary disability. I use voice when I'm on mobile. I can describe the problem, get a program generated, click run to verify it.

3. Maybe. Their feature rollouts feel more like what other companies do via unannounced A/B testing.

Re: We are beginning to roll out new voice and image capabilities in ChatGPT

#440

Earlier quoted context omitted.

You can see the difference if you know where to poke. For instance, if you start making spatial abstractions ChatGPT will often make mistakes, you can point it out, they can explain why it's a mistake, but it has no internalized model of what these words mean, so it keeps making the same mistakes (see here for a better idea of what I'm talking about[1]). The fact that you are interacting with it through text means th…

This is also true of humans. Many school students will hands in answers they don't understand in the hope of getting the mark and then try to cover themselves when asked about it, even if they repeat the same mistakes.

Trying to make things up to cover for a lack of knowledge is something distinctly different, though. This is a a situation where ChatGPT is able to perfectly describe the mistake it made, describe exactly what it needs to do differently, and then keeps making the same mistake, even with simple tasks. That’s because there’s no greater model that the words are being connected to.

The equivalence would be saying to someone, “put this on the red plate, not the blue one.” And they say sure, then put it on the blue one. You tell them they made a mistake and ask them if they know what it was, and they reply “I put it on the blue plate, not the red one. I should have put it on the red one.” Then you ask them to do it again, and they put it on the blue plate again. You tell them no, you made the same mistake, put it on the blue plate, not the red one. They reply with, “Sorry, I shouldn’t have put it on the blue plate again, now I’m going to put it on the red one,” and then they put it on the blue plate yet again.

Do humans make mistakes? Sure. But that kind of performance in a test wouldn’t be considered a normal mistake, but rather a sign of a serious cognitive impairment.

Post reply on HN