We are beginning to roll out new voice and image capabilities in ChatGPT
241–250 of 914 posts
Re: We are beginning to roll out new voice and image capabilities in ChatGPT
#242The thought of my children being put to bed by a machine is horrifying. Then again, perhaps this is better than many kids have. Shudder.
If I could harness the power of AI to outsource my tasks, reading bedtime stories to my kids would be the last thing on that list. That's cherished time. Those are lifelong memories. Those are the moments we are supposed to be striving to have more of. It saddens me to think of the amount of engineering work that went into creating that example while entirely missing the point. These are the moments we are supposed t…
We have major priority issues from what I can see. If we want to live our lives more but put an AI to work doing something we tend to claim we place very high in our value hierarchy, we’re effectively inviting death into life. We’re forfeiting something we love. That’s incredibly sad to me.
Re: We are beginning to roll out new voice and image capabilities in ChatGPT
#243Re: We are beginning to roll out new voice and image capabilities in ChatGPT
#244This announcement seem to have killed so many startups that were trying to do multi-modal on top of ChatGPT. The way it's progressing with solving use cases with images and voice, not too far when it might be the 'one app to rule them all'. I can already see "Alexa/Siri/Google Home" replacement, "Google Image Search" replacement, ed-tech startups that were solving problems with AI using by taking a photo are also doo…
1. Domain-specific AI - Training an AI model on highly technical and specific topics that general-purpose AI models don't excel at.
2. Integration - If you're going to build on an existing AI model, don't focus on adding more capabilities. Instead, focus on integrating it into companies' and users' existing workflows. Use it to automate internal processes and connect systems in ways that weren't previously possible. This adds a lot of value and isn't something that companies developing AI models are liable to do themselves.
The two will often go hand-in-hand.
Re: We are beginning to roll out new voice and image capabilities in ChatGPT
#245My biggest complaint with OpenAI/ChatGPT is their horrible "marketing" (for lack of a better term). They announce stuff like this (or like plugins), I get excited, I go to use it, it hasn't rolled out to me yet (which is frustrating as a paying customer), and my only recourse is.... check back daily? They never send an email "Plugins are available for you!", "Voice chat is now enabled on your account!" and so often I…
Re: We are beginning to roll out new voice and image capabilities in ChatGPT
#246Earlier quoted context omitted.
In retrospect, such startups should have been wary: they should have known that OpenAI had Whisper, and also that GPT-4 was designed with image modality. I wouldn't say that OpenAI "telegraphed" their intentions, but the very first strategic question should have been, "Why isn't OpenAI doing this already, and what do we do if they decide to start?"
>I wouldn't say that OpenAI "telegraphed" their intentions They did telegraph it, they showed the multimodal capabilities back in the GPT4 Developer Livestream[0] right before first releasing it. 0. https://youtu.be/outcGtbnMuQ?t=943
I think the only place where plugins will make sense are for realtime things like booking travel or searching for sports/stock market/etc type information.
Re: We are beginning to roll out new voice and image capabilities in ChatGPT
#247Yet it still can't tell me how to import the Redirect type from Next.js and lies about it.
Re: We are beginning to roll out new voice and image capabilities in ChatGPT
#248This announcement seem to have killed so many startups that were trying to do multi-modal on top of ChatGPT. The way it's progressing with solving use cases with images and voice, not too far when it might be the 'one app to rule them all'. I can already see "Alexa/Siri/Google Home" replacement, "Google Image Search" replacement, ed-tech startups that were solving problems with AI using by taking a photo are also doo…
Talking to Google and Siri has been positively frustrating this year. On long solo drives, I just want to have a conversation to learn about random things. I've been itching to "talk" to chatGPT and learn more (french | music theory | history | math | whatever) all summer. This should hit the spot!
Example from a couple days ago:
Me, in the shower so not able to type: "Hey Siri, add 1.5 inch brad nails to my latest shopping list note."
Siri: "Sorry, I can't help with that."
... Really, Siri? You can't do something as simple as add a line to a note in the first-party Apple Notes app?
Re: We are beginning to roll out new voice and image capabilities in ChatGPT
#249My biggest complaint with OpenAI/ChatGPT is their horrible "marketing" (for lack of a better term). They announce stuff like this (or like plugins), I get excited, I go to use it, it hasn't rolled out to me yet (which is frustrating as a paying customer), and my only recourse is.... check back daily? They never send an email "Plugins are available for you!", "Voice chat is now enabled on your account!" and so often I…
Re: We are beginning to roll out new voice and image capabilities in ChatGPT
#250This is the dagger that will make online schooling unviable. ChatGPT already made it so that you could easily copy & paste any full-text questions and receive an answer with 90% accuracy. The only flaw was that problems that also used diagrams or figures would be out of the domain of ChatGPT. With image support, students could just take screenshots or document scans and have ChatGPT give them a valid answer. From wha…
This is obviously not easy or going to happen without time and resources, but that is how adaptation goes.