Earlier quoted context omitted.
What do you mean? Everyone has access to the gpt-4o model right now through ChatGPT and the API. Sure we don't have voice-to-voice but we have a lot more than what Google has promised.
API yes, ChatGPT no (at least not for all users); I've got my own web interface for the API so I can play with the model (for all of $0.045 of API fees), but most people can't be bothered with that and will only get 4o when it rolls out as far as their specific ChatGPT account.
Veo
161–170 of 539 posts
Re: Veo
#162Earlier quoted context omitted.
What do you mean? Everyone has access to the gpt-4o model right now through ChatGPT and the API. Sure we don't have voice-to-voice but we have a lot more than what Google has promised.
API yes, ChatGPT no (at least not for all users); I've got my own web interface for the API so I can play with the model (for all of $0.045 of API fees), but most people can't be bothered with that and will only get 4o when it rolls out as far as their specific ChatGPT account.
The bigger issue is that 4o without the multi-modal, new speech capabilities or desktop app isn't that different to GPT-4. And those things aren't yet launched.
Re: Veo
#163Earlier quoted context omitted.
"Laser-focused on the bottom line at the expense of all else" is not how I'd describe Google, now or at any point in the past. They have a lot of dysfunction, but if anything that dysfunction stems from too much experimentation and autonomy at the leaf nodes of the organization. That's how they get into these crazy places where they have to pick between 5 chat apps or whatever. If Google were as focused on ads as you…
I'd describe Google as focused on the bottom line after they put the ads guy in charge of search. I'm referring to this article that was posted here recently: https://www.wheresyoured.at/the-men-who-killed-google/
Re: Veo
#164Not nearly as impressive as Sora. Sora was impressive because the clips were long and had lots of rapid movement since video models tend to fall apart when the movement isn't easy to predict. By comparison, the shots here are only a few seconds long and almost all look like slow motion or slow panning shots cherrypicked because they don't have that much movement. Compare that to Sora's videos of people walking in rea…
Re: Veo
#165Not nearly as impressive as Sora. Sora was impressive because the clips were long and had lots of rapid movement since video models tend to fall apart when the movement isn't easy to predict. By comparison, the shots here are only a few seconds long and almost all look like slow motion or slow panning shots cherrypicked because they don't have that much movement. Compare that to Sora's videos of people walking in rea…
I think this is arguably better than the alternative. With slow-mo generated videos, you can always speed them up in editing. It's much harder to take a fast-paced video and slow it down without terrible loss in quality.
Re: Veo
#166Sigh. I still have hopes for VEO though
Re: Veo
#167> Veo > Sign up to try VisionFX Is it Veo or VisionFX? Is it a sign up, a trial, or a waitlist? How hard can it be to write a clear message? In the words of Don Miller, if you confuse, you lose.
Disclaimer: I work at Google on related stuff Veo is the name of a video model. VideoFX is the name of a new experimental tool at labs.google.com, which uses Veo and lets you make videos. Thanks for the feedback though, I see how it's confusing for users.
Re: Veo
#168Re: Veo
#169Not nearly as impressive as Sora. Sora was impressive because the clips were long and had lots of rapid movement since video models tend to fall apart when the movement isn't easy to predict. By comparison, the shots here are only a few seconds long and almost all look like slow motion or slow panning shots cherrypicked because they don't have that much movement. Compare that to Sora's videos of people walking in rea…
being cautious often puts a dent in innovation
Re: Veo
#170Earlier quoted context omitted.
What do you mean? Everyone has access to the gpt-4o model right now through ChatGPT and the API. Sure we don't have voice-to-voice but we have a lot more than what Google has promised.
How do I get access? I just checked my app and the Premium upgrade says it will unlocked GPT-3.5 and GPT-4, so I assume my version is still the old one. All my apps are updated in the App Store too.