Live data from Hacker News

GPT‑Live

openai.com

241–250 of 556 posts

Re: GPT‑Live

#241
post #19

I had preview access to this one for a few weeks. It's very good. I had one conversation that lasted a full hour while I was walking the dog, got some good brainstorming done against one of my projects. The best feature is that it can delegate questions out to GPT-5.5 in the background, so you're no longer restricted to a voice model that's several years behind the frontier. I did report a fun bug with it though: it…

I highly recommend simply enjoying the walk.

I take my dog for walks everyday and I think she 100% of the times enjoys it but me maybe only 20-30%

Re: GPT‑Live

#242
post #211

Earlier quoted context omitted.

Hey! Bit of an unusual question maybe: if this stuff further exarcerbates the loneliness epidemic and atomization of society, will you be able to live with yourself you think? If you hear about teenagers only spending time with your chatbot in 5 years, will you feel some amount of personal responsibility or not? Always curious to hear you guys' perspective on that kind of stuff!

"Hey Leibniz, how do you live with yourself knowing that your binary system helped eventually replace human conversations?"

[deleted]

Re: GPT‑Live

#243
post #211

Earlier quoted context omitted.

Hey! Bit of an unusual question maybe: if this stuff further exarcerbates the loneliness epidemic and atomization of society, will you be able to live with yourself you think? If you hear about teenagers only spending time with your chatbot in 5 years, will you feel some amount of personal responsibility or not? Always curious to hear you guys' perspective on that kind of stuff!

"Hey Leibniz, how do you live with yourself knowing that your binary system helped eventually replace human conversations?"

Fair question, although I think he really has a hard time living with it ...

Re: GPT‑Live

#244

This is the opposite direction AI should be going. Human relationships are the most valuable thing we have, and so, naturally, technology seeks to intermediate and now replace them. I'm not Catholic, but this podcast presents a very interesting argument against talking to AI as if they were human: https://newpolity.com/podcasts-hub/debate-chatbots

How would talking to an AI as if it were not human sound? You can probably set your system prompt to insert “beep boop” between sentences and make it refer to itself as “Cybertron9000 Personal Computing Device” if that’s what you like. Is that an improvement? Or are you against voice computer interfaces altogether?

My preference would be to turn down the fake emotional expressions and notes in the responses. No cheeky quips, etc.

Otherwise it's kind of like being manipulated by a psychopath

Re: GPT‑Live

#245
post #212

Earlier quoted context omitted.

Wait, you can talk to Teslas now? How did I miss thiS? Can I get a red led bar and basically have a KITT?

the cybertruck seems designed for this KITT fantasy of yours

Dammit! I live in Europe and the cyber truck isn't even available here..

That said a huge pickup truck is about as far as you can get from a Camaro... Then again I'm not exactly David Hasslehoff myself either... Meh if it talks that's close enough!

Re: GPT‑Live

#246

(Atty from OpenAI here) GPT-Live-1 is the first version of a new generation of models, and we believe the full-duplex architecture + delegation enables entirely new ways of human-AI interaction. Would love to hear your feedback!

Hey! Bit of an unusual question maybe: if this stuff further exarcerbates the loneliness epidemic and atomization of society, will you be able to live with yourself you think? If you hear about teenagers only spending time with your chatbot in 5 years, will you feel some amount of personal responsibility or not? Always curious to hear you guys' perspective on that kind of stuff!

I love the chatty tone of this utterly dystopian question.

Re: GPT‑Live

#247
post #19

I had preview access to this one for a few weeks. It's very good. I had one conversation that lasted a full hour while I was walking the dog, got some good brainstorming done against one of my projects. The best feature is that it can delegate questions out to GPT-5.5 in the background, so you're no longer restricted to a voice model that's several years behind the frontier. I did report a fun bug with it though: it…

> "The best feature is that it can delegate questions out to GPT-5.5 in the background, so you're no longer restricted to a voice model that's several years behind the frontier." Wow. That's exactly what I hoped they would do. This issue has held me back from using ChatGPT's voice mode as much as I otherwise would have, because I also use it for brainstorming while commuting, exercising, etc., and don't want it to fe…

Funnily enough - I built this (delegation) over the weekend with Fable for a local voice chat running 100% on local LLMs, Parakeet and Kokoro. I say "...ask the thinking model..." and that redirects it to Qwen 3.6 27B on vLLM.

Can't claim originality though - it was inspired by Sesame - where their models will invoke a search, or check the weather etc, and make a vocalisation to keep you engaged.

Turn taking is one of the hardest things to get right for the exact reasons mentioned - but does seem to be the way that Claude.ai's voice works - in a very obvious way.

Anthropic + OpenAI both rug-pulled voices I liked and got used to and OpenAI really dumbed down their voices at the same time - Arbor went from Estuary English and almost "jack the lad" to some generic English accent. Claude had a Birmingham accent and said things like "shit", ending sentences like "So you're telling me that they asked for a 90% discount yeah?" - then it changed overnight to a mock Derbyshire accent with a dull tone.

ChatGPT's voice also gaslights me for conventional opinions - "my Eastern European neighbour helped me lift a wardrobe upstairs - something you just can't ask your typical neighbour neighbour"... then you get a full on left-leaning lecture from the safety layers rather than a head nod or "what luck!"

Claude + Sesame are nowhere near as overbearing.

In both cases - from edgy and engaging to something that just didn't gel.

The point of making my own assistant is that I can talk for as long as I want, episodic memory is personal and private, there's no "trust me bro, we're a big corporation" vibes.

This was not my first attempt - when I had a bunch of Opus credit around Jan/Feb - I tried really hard and created something that was not good enough. What I have now, is working, and each session is training Claude/Codex on what to tune, and to fix.

"Just had a convo - can you look into what happened?" And if it's one I don't mind sharing with the model - I'll say, "and what did you think of the questions I asked?" Sometimes it'll give a lovely commentary on how the model did.

af_heart is probably the smoothest voice - but yes it's more like another commented - more "StarTrek" than "telesales assistant that pauses and laughs at your jokes".

If you're on a similar path and want something full duplex - the go to solution is PersonaPlex from Nvidia based upon Moshi.

Re: GPT‑Live

#248

This is the opposite direction AI should be going. Human relationships are the most valuable thing we have, and so, naturally, technology seeks to intermediate and now replace them. I'm not Catholic, but this podcast presents a very interesting argument against talking to AI as if they were human: https://newpolity.com/podcasts-hub/debate-chatbots

Isn’t voice I difference in degree rather than in kind? I definitely talk to AI as if it were human (one might say the UX of AI is to emulate a human). And a large portion of my interaction with humans is via text, for example, this post!

> Isn’t voice I difference in degree rather than in kind?

There is a difference in expression / emotionality with speaking vs writing. Speaking tends to carry more emotion while writing is generally more deliberate/less-emotional*. Having a voice conversation will be more likely to get a human to engage in an emotional based expression mode, which could increase the chance of "false connections", believing the AI "gets" them or "understands" or "listens". This happens with text too, as some headlines show.

The issue is that while some people are going to "connect" with their AI in text and voice, some who do not make the connection via text may do so via voice because it tends to change a persons expression mode.

> I definitely talk to AI as if it were human

Do you talk to it as if it were a friend or family? or Do you just use natural language to give directives? The distinction, I believe, is in the kind of way we express the "talking".

I talk to AI as if it were a tool that understands human commands and then executes those commands and relays them in a human understandable format. This includes commands to provide options that I may not have covered and explain the options. If I talked to a human this way, they wouldn't be around much longer -- unless they were an employee and even then they would probably be looking for a new job

After reading some of the psychotic break headlines from AI chats, I see some people really do talk to AI as if it were human. Which I would guess includes seeking broad "thoughts and feelings" on a persons situation or asking the AI if their view/side of things is the "right or wrong" side. Basically begging the AI to be responsible for their own thoughts, or simply offloading them and taking what comes back -- which is going to be what they wanted to hear because the entire context would be full of emotion based prompts.

*I forget which books Ive read about this in. It's not an obscure concept, quick search brought this up: https://kellercenter.hankamer.baylor.edu/news/story/2023/spe...

Re: GPT‑Live

#249

I reaaaaaally hope we have an option to disable those random ums and ahhs that interrupt for no reason. :|

I was surprised all that made it into the demo video. Did not seem cool. Is that "active listening" or something?

Its actually useful, when you are launching into a long monologue and want periodic acknowledgement that its "listening".

Re: GPT‑Live

#250
post #222

Earlier quoted context omitted.

I highly recommend simply enjoying the walk.

Given the personality type common on HN, I imagine that the GP, even if unplugged from all technology on their walk, wouldn't be in a mindful state of enjoying their surroundings, but rather would be "lost in the clouds", stewing on the same ideas/thoughts/problems; but with those thoughts going more in circles, due to a lack of ability to verify anything.

Literally me. Everyone is different, and that's fine. But I don't have the privilege of living in an area where I can talk to people about the things I am thinking through. It's very rural. Having a _utility_ that can act as a sounding board while I spew out my thoughts on a walk is a really meaningful improvement to my current situation.
Post reply on HN