Live data from Hacker News

GPT‑Live

openai.com

281–290 of 556 posts

Re: GPT‑Live

#281

This voice is awful , possibly one of the worst AI/computer voices I have heard in like two years now - what's up with that? Is this seriously the best a company burning this much money can do, and they consider this acceptable to release? Like two syllables in and my first response was to grimace and physically cringe. Does anyone here think this sounds good, or even just "fine enough"? Do engineers that work on thi…

What don't you like about the voice? Geniune question, I'm not using it in English but I think it sounds fine in my language, and it's an improvement over the previous one.

Re: GPT‑Live

#282
If an idiot has this on in the subway my conversations are surveilled. What is the antidote? Train another model to talk about bombs etc. and flood the clanker (and by extension the FBI)?

Re: GPT‑Live

#283

Feel like the intro video is very odd. Basically have an older lady (not their target audience) blatantly reading a teleprompter. Why are they going after this audience? Retired people have no use for delegated tasks or information. They also are the least likely to use it and not get frustrated.

GPT-Live... buy me that thing I saw on tv

Re: GPT‑Live

#285
post #19

I had preview access to this one for a few weeks. It's very good. I had one conversation that lasted a full hour while I was walking the dog, got some good brainstorming done against one of my projects. The best feature is that it can delegate questions out to GPT-5.5 in the background, so you're no longer restricted to a voice model that's several years behind the frontier. I did report a fun bug with it though: it…

I highly recommend simply enjoying the walk.

One thing has remained constant over the last little period of time. AI boosters have zero taste.

Re: GPT‑Live

#286
I have not used voice mode much with chatgpt. I was surprised to learn that they were already not running the voice model like a UX orchestrator while utilizing other models in background for actual research/response etc. I guess it's good they launched what they could and got here in steps. I suspect in the near future my personal device (mobile/laptop) will be powerful enough to run any UX orchestrator model locally – and route to multiple frontier closed/open model providers in the background as appropriate. The battle is going to be platform owners (Apple/Google/Microsoft) wanting to lock-down the access to that local hardware and local interaction paradigms (ambient always-on full-duplex voice) and intermediate through their platform layers - rationalizing it as consumer security/privacy protection (which is right for most people, but sucks for the open market). Meanwhile I suspect OpenAI/Meta et al will try to build their own hardware and become platform owners themselves, though unsuccessfully. And it's going to take some company like epic games to get them to open that up. and that's probably what the next decade is going to be all about.

Re: GPT‑Live

#287

Earlier quoted context omitted.

For many years, I've wanted ED-209 (robocop) voice from something like espeak or similar. Still can't find anything good. Not for chat, just as a way to make notification messages that sound like ED-209.

One popular speech synth from back in the day, I believe it was WillowTalk, had a voice called Colossus, which sounded like the voice module of the computer from Colossus: The Forbin Project . This voice was used for that of CATS in the famous "All your base are belong to us" Flash video. Another WillowTalk voice was a clone of DECtalk's Perfect Paul good enough to be used as the voice in the MC Hawking rap recording…

I am having a hell of a time googling for WillowTalk, do you know any links/places where I can learn more? Maybe even download something?

Re: GPT‑Live

#288

Earlier quoted context omitted.

I was surprised all that made it into the demo video. Did not seem cool. Is that "active listening" or something?

Its actually useful, when you are launching into a long monologue and want periodic acknowledgement that its "listening".

I think it would make me stop speaking. I guess it might take me some time to get used to it.

Re: GPT‑Live

#289
post #19

I had preview access to this one for a few weeks. It's very good. I had one conversation that lasted a full hour while I was walking the dog, got some good brainstorming done against one of my projects. The best feature is that it can delegate questions out to GPT-5.5 in the background, so you're no longer restricted to a voice model that's several years behind the frontier. I did report a fun bug with it though: it…

>The best feature is that it can delegate questions out to GPT-5.5 in the background, so you're no longer restricted to a voice model that's several years behind the frontier.

Ahh, this makes sense. I was wondering when they would start doing this. I stopped using voice mode all together because it was frustrating talking to a dumb AI, when most of the time I discuss things with Opus 4.8 or gpt 5.5.

I was working on a phone call agent recently, and thought about doing this. It makes sense

Re: GPT‑Live

#290
I worry what this will do to human communication if it becomes commonplace. Will everyone learn to be a forceful speaker, speaking over anyone they want to stop speaking?
Post reply on HN