Live data from Hacker News

We are beginning to roll out new voice and image capabilities in ChatGPT

openai.com

811–820 of 914 posts

Re: We are beginning to roll out new voice and image capabilities in ChatGPT

#811

Earlier quoted context omitted.

Your username checks out! That said, is it that much different from the past twenty years, when everyone was being told to follow their passion and get a useless $200,000 communication or literature degree to then go work at Starbucks? At least kids growing up with AI will have a chance to make its use second nature like many of us did with computers 20-30 years ago. The kids with poor parental/counselor guidance wil…

I do think it is much different from the past twenty years. Twenty years ago we didn't have ChatGPT. There are things we could compare it to, but there also isn't anything like it. My biggest fear is just a lack of jobs. When people need experience to work, and the work you give to people to give them experience is replaced by ChatGPT - then what do we do? Of course there will still be companies hiring people, but wh…

The only way there will be no jobs is if every conceivable human need is met by robots. In which case there will also be no need to work.

Re: We are beginning to roll out new voice and image capabilities in ChatGPT

#812

Earlier quoted context omitted.

If you focus on integration, you're up against autogpt, gorilla, etc.

AutoGPT isn't remotely usable for practical enterprise software purposes right now.

Agreed, not yet. It will get there.

Re: We are beginning to roll out new voice and image capabilities in ChatGPT

#813

Voice has the potential to be awesome. This demo is really underwhelming to me because of the multi-second latency between the query and response, just like every other lame voice assistant. It doesn't have to be this way! I have a local demo using Llama 2 that responds in about half a second and it feels like talking to an actual person instead of like Siri or something. I really should package it up so people can t…

Completely agree, latency is key for unlocking great voice experiences. Here's a quick demo I'm working on for voice ordering https://youtu.be/WfvLIEHwiyo Total end-to-end latency is a few hundred milliseconds: starting from speech to text, to the LLM, then to a POS to validate the SKU (no hallucinations are possible!), and finally back to generated speech. The latency is starting to feel really natural. Building out…

Nice work, very cool!

Re: We are beginning to roll out new voice and image capabilities in ChatGPT

#814
post #795

Earlier quoted context omitted.

You have no clue how GPT-4 functions so I don't know why you're assuming they're "thinking in language"

I am comfortable asserting that an LLM like GPT-4 is only capable of thinking in language; there is no distinction for an LLM between what it can conceive of and what it can express.

It certainly "thinks" in vector spaces at least. It also is multimodal, so not sure how that plays in?

Re: We are beginning to roll out new voice and image capabilities in ChatGPT

#815

I keep hoping to be able to give it a jpg of handwritten text and it'll give me back ASCII text.

This... would be amazing. Handwritten OCR has been hit or miss, requiring a collection of penstroke data for most recognizers to work, and they work poorly at that.

Re: We are beginning to roll out new voice and image capabilities in ChatGPT

#816

Earlier quoted context omitted.

My argument was that improvement from scale would continue. There is absolutely evidence suggesting this. Gpt-4 can perform nearly all tasks you throw at it with well above average human performance. There literally isn't any testable definition of intelligence it fails that a big chunks of humans wouldn't also fail. You seem to keep missing the fact that We do not need an exponential improvement from 4.

> Gpt-4 can perform nearly all tasks you throw at it with well above average human performance. It can't even generate flashcards from a textbook chapter, because it can't load the entire chapter into memory. Heck, it doesn't even know what textbook I'm talking about; I have to provide the content! It fails constantly at real world coding problems, and often does so silently. If you tried to replace a software develo…

>It can't even generate flashcards from a textbook chapter, because it can't load the entire chapter into memory. Heck, it doesn't even know what textbook I'm talking about; I have to provide the content!

Okay...? That's a context window problem. and you could manage it if you sent the textbook in chunks.

>The improvement GPT 5 would have to provide is multiple orders of magnitude in order for this to be a realistic proposition.

No..it wouldn't

https://arxiv.org/abs/2309.12499

Re: We are beginning to roll out new voice and image capabilities in ChatGPT

#817

Earlier quoted context omitted.

I read this all the time and yet no one can seem to come up with even a few questions from several months ago that ChatGPT has become “worse” at. You would think if this is happening it would be very easy to produce such evidence since chat history of all conversations is stored by default.

Everytime it’s mentioned someone says this and other users provide examples. Maybe you just don’t care about those examples

Care to share these examples, in a scientific (n > 30) manner that can’t just be attributed to model nondeterminism? I don’t follow these threads religiously but in the ones I’ve seen no one has been able to provide any sort of convincing evidence. I’m not some sort of OpenAI apologist, so if there is actual good provable evidence here I will easily change my mind about it

Re: We are beginning to roll out new voice and image capabilities in ChatGPT

#818

Earlier quoted context omitted.

I read this all the time and yet no one can seem to come up with even a few questions from several months ago that ChatGPT has become “worse” at. You would think if this is happening it would be very easy to produce such evidence since chat history of all conversations is stored by default.

Here's a specific example https://news.ycombinator.com/item?id=37533417

Pointing out a specific bug with functionality is not the same as saying “in general the quality of GPT answers has decreased over X months” especially when that bug is in a realm that LLM’s have already been provably bad at.

Re: We are beginning to roll out new voice and image capabilities in ChatGPT

#819
post #58

Earlier quoted context omitted.

I read this all the time and yet no one can seem to come up with even a few questions from several months ago that ChatGPT has become “worse” at. You would think if this is happening it would be very easy to produce such evidence since chat history of all conversations is stored by default.

OpenAI regularly changes the model and they admit the new models are more restricted, in the sense that they prevent tricky prompts from producing naughty words, etc. It should be their responsibility to prove that it's just as capable.

He who makes the logical argument must provide the burden of proof. Did OpenAI claim that their models didn’t regress while putting these new safeguards into place? If not, it feels like the burden of proof lies on whoever said that they did.

To be specific, the claim we are talking about here is “ChatGPT gives generally worse answers to the exact same questions than ChatGPT gave X months ago”. Perhaps for the subset of knowledge space you reference that updates were pushed to that is pretty easily provably true, but I’m more interested in the general case.

In other words, you can pretty easily make the claim that ChatGPT got worse at telling me how to make a weapon than it did 3 months ago. I could pretty easily believe that and also accept that it was probably intentional. While we can debate whether it was a good idea or not, I’m more interested in the claim over whether ChatGPT got worse at summarizing some famous novel or helping write a presentation than it was 3 months ago.

Re: We are beginning to roll out new voice and image capabilities in ChatGPT

#820

Earlier quoted context omitted.

I read this all the time and yet no one can seem to come up with even a few questions from several months ago that ChatGPT has become “worse” at. You would think if this is happening it would be very easy to produce such evidence since chat history of all conversations is stored by default.

> I read this all the time and yet no one can seem to come up with even a few questions from several months ago that ChatGPT has become “worse” at this could just mean that people do not have time to argue with strangers

Well, sure, but shouldn’t some pedant have the time to dig up their ChatGPT history from 4 months ago to disprove the claim? Seems like it would be pretty easy to do and there are plenty of pedants on the internet but I don’t see the blogosphere awash of side by side comparisons showing how much worse it got
Post reply on HN