Earlier quoted context omitted.
This is pretty good. Do you think running models locally will be able to achieve performance (getting task done successfully) compared to cloud based ones.i am assuming for context of a drive through scenario it should be ok but more complex systems might need external infromation
Definitely depends on the application, agreed. The more open ended the application the more dependent it is on larger LLMs (and other systems) that don't easily fit on edge. At the same time, progress is happening that is increasing the size of LLM that can be ran on edge. I imagine we end up in a hybrid world for many applications, where local models take a first pass (and also handle speech transcription) and only…
We are beginning to roll out new voice and image capabilities in ChatGPT
691–700 of 914 posts
Re: We are beginning to roll out new voice and image capabilities in ChatGPT
#692Earlier quoted context omitted.
Firstly no, the gap between 3 and 4 is not anything as large as the gap between 2 and 3. Secondly, nothing you said here changed as of this announcement. Nothing here makes it any more or less likely LLMs will risk software engineering jobs. Thirdly, you can take what Sam Altman says with as many grains of salt as you like, if there really was no innovation at all as you claim, then there will be a limit hit at compu…
>the gap between 3 and 4 is not anything as large as the gap between 2 and 3. We'll just have to agree to disagree. 3 was a signal of things to come but it was ultimately a bit of a toy, a research curiosity. Utility wise, they are worlds apart. >if there really was no innovation at all as you claim, then there will be a limit hit at computing capability and cost. computing capability and cost are just about the one…
And it is not true that computing power will continue to reduce; Moore's Law has been dead for some time now, and if incremental growth in LLMs require exponential growth in computing power the marginal difference won't matter. You would need a matching exponential growth in processing capability which is most certainly not occurring. So compute will not fall at the rate you would need it to for LLMs to actually compete in any meaningful way with human software engineers.
We are not guaranteed to continue to progress in anything just because we have in the past.
Re: We are beginning to roll out new voice and image capabilities in ChatGPT
#693Earlier quoted context omitted.
Talking to Google and Siri has been positively frustrating this year. On long solo drives, I just want to have a conversation to learn about random things. I've been itching to "talk" to chatGPT and learn more (french | music theory | history | math | whatever) all summer. This should hit the spot!
It’s funny. Driving buddy has been my number one use case for a while now. Still can’t quite make it work. I feel like I could learn a lot if I could have random conversations with GPT. + bonus if someone else in the car got excited when I see cows. Don’t care if it’s an AI.
Re: We are beginning to roll out new voice and image capabilities in ChatGPT
#694Earlier quoted context omitted.
>, the "mental" workers who get replaced by AI could simply move to manual jobs and total production and average wages would go up Why would manual job average wages go up? You're increasing the size of the labor pool.
Total production would increase (AI will allow us to make more with less) and I'm expecting the capital / labor share to remain stable. An analogy: Imagine that half of the labor force makes cars, the other half creates software. The average person buys 1 car and 1 software per year. There's a breakthrough, AI can now be used to create software almost for free. It can even make 2x more software per year. The programm…
I don't have to argue.. others have done it for me
https://www.cnbc.com/2022/04/01/richest-one-percent-gained-t...
https://www.cnbc.com/2023/01/16/richest-1percent-amassed-alm...
https://time.com/5888024/50-trillion-income-inequality-ameri...
Re: We are beginning to roll out new voice and image capabilities in ChatGPT
#695Earlier quoted context omitted.
Completely agree, latency is key for unlocking great voice experiences. Here's a quick demo I'm working on for voice ordering https://youtu.be/WfvLIEHwiyo Total end-to-end latency is a few hundred milliseconds: starting from speech to text, to the LLM, then to a POS to validate the SKU (no hallucinations are possible!), and finally back to generated speech. The latency is starting to feel really natural. Building out…
Manna v0.7
Re: We are beginning to roll out new voice and image capabilities in ChatGPT
#696I went from being worried to thinking it won't replace me anytime soon after using GPT4 for a while and now I'm back to being worried. Because the pace of development is intense. I would love to be financially independent and watch this with excitement and perhaps take on risky and fun projects. Now I'm thinking - how do I double or triple my income so that I reach financial independence in 3 years instead of 10 year…
I'm very worried constantly. This is the story of the bear, where you just have to be faster than the other guy. For now. The bear is getting faster and faster and it won't be long before it eats all of us. It feels like we're at the end of history. I don't know where we go from here but what are we useful for once this thing is stuck inside a robot like what Tesla is building? What is the point of humanity? Even tak…
Re: We are beginning to roll out new voice and image capabilities in ChatGPT
#697We should be fine as long as it doesn't move. Jokes aside, I have paused my subscription because even GPT4 seemed to become dumber at tasks to the point that I barely used it, but the constant influx of new features is tempting me to renew it just to check them out...
I switched to Claude, it's better at explaining stuff in a more direct manner without the always-excited way of talking. Is that an engagement trick? Maybe ChatGPT is intended to be more of a chatbot that you can share your thoughts with.
I don't agree with this perspective. These aren't rigid systems that only respond one way. If you want it to respond a certain way, tell it to.
This is the purpose of custom instructions, in ChatGPT, so you only have to type the description once.
Here's mine, modeled on a few I've seen mentioned here:
You should act as an expert.
Be direct.
Do not offer unprompted advice or clarifications.
Never apologize.
And, now there's support for describing yourself to it. I've made it assume that I don't need to be babied, with the following puffery: Polymath. Inquisitive. Abstract thinker. Phd.
Making it get right into the gritty technicalities.edit: or, have it respond as a grouchy space cowboy, if you want.
Re: We are beginning to roll out new voice and image capabilities in ChatGPT
#698Earlier quoted context omitted.
I'm not convinced that this pace will continue. We're seeing a lot of really cool, rapid evolution of this tech in a short amount of time, but I do think we'll hit a soft ceiling in the not too distant future as well. If you look at something like smartphones, for example. Smartphones, from my perspective, got drastically better and better from about ~2006-2015 or so. They were rapidly improving cameras and battery l…
> I do think we'll hit a soft ceiling in the not too distant future ... it's going to plateau and progress will become substantially more gradual. I don't think this will age well. It's a matter of simple compute power to advance from realistic text/token prediction, to realistic synthesis of stuff like human (or animal) body movement, for all kinds of situations, including realistic facial/body language, moods, and…
Yup. It's "just" a compute advance away. Never mind it's already consuming as much computing as we can throw at it. It's "just" there.
Re: We are beginning to roll out new voice and image capabilities in ChatGPT
#699Earlier quoted context omitted.
I'm very worried constantly. This is the story of the bear, where you just have to be faster than the other guy. For now. The bear is getting faster and faster and it won't be long before it eats all of us. It feels like we're at the end of history. I don't know where we go from here but what are we useful for once this thing is stuck inside a robot like what Tesla is building? What is the point of humanity? Even tak…
re: UBI. I don't think they'll let us starve, but that's a very low bar. If we all become fungible and invaluable they can just feed us Soylent green.
Re: We are beginning to roll out new voice and image capabilities in ChatGPT
#700Earlier quoted context omitted.
"londons_explore" - Ahh the classic British cynicism (Don't ban-ish me señor Dang, I'm British so I can say this). > Similar possibilities existed in medicine for 50 years It would've been like building the tower of babel with a bunch of raspbery pi zeros. While theoretically possible, practically impossible and not (just) because of laws, but rather because of structural limitations (vector dbs of the internet solve…
> I'm British so I can say this can you, though? it's not scalably confirmable. what you can say in a British accent to another human person in the physical world is not necessarily what you can say in unaccented text on the internet.
Funnily enough, it is scalably confirmable. You can feed all my HN comments before chatGPT into well.. chatGPT and ask it whether I'm british based on the writing.
I bet we are just a version or two away from being able fine tune it down to region based on writing. There are so many little things based on whether your from Scotland, Wales or London. Especially London!