Live data from Hacker News

Veo

deepmind.google

241–250 of 539 posts

Re: Veo

#241
post #97
post #85

Earlier quoted context omitted.

What do you mean? Everyone has access to the gpt-4o model right now through ChatGPT and the API. Sure we don't have voice-to-voice but we have a lot more than what Google has promised.

How do I get access? I just checked my app and the Premium upgrade says it will unlocked GPT-3.5 and GPT-4, so I assume my version is still the old one. All my apps are updated in the App Store too.

In the App Store there's a new build of the iOS app as of 3 hours about (call it about 11am US Pacific time). It includes the GPT-4o model (at least it shows it for me.)

Re: Veo

#242

all of this stuff i'll believe when it's ready for public release 1. safety measures lead to huge quality reductions 2. the devil's in the details. you can make me 1 million videos which look 99% realistic, but it's useless. consumers can pick it instantly, and it's a gigantic turn-off for any brand

[deleted]

Re: Veo

#243

Earlier quoted context omitted.

Is Sora available in any country?

I thought I read they’ve deemed Sora too dangerous to release pre- election ? Or have reservations about it ? I might be wrong…

Sounds like a great excuse / communications strategy!

Re: Veo

#244

> Veo > Sign up to try VisionFX Is it Veo or VisionFX? Is it a sign up, a trial, or a waitlist? How hard can it be to write a clear message? In the words of Don Miller, if you confuse, you lose.

This is very on-brand with how Google does branding. "are you confused yet? no? try this other vaguely similar name."

Maybe it's going to be a new messaging app - but with AI!

Kidding... I signed up for the waitlist. I have ideas for videos I'd like to use to explain things that I have no hope of creating myself.

Re: Veo

#245

With so much recent focus by OpenAI/Google on AI's visual capabilities, does anyone know when we might see an OCR product as good as Whisper for voice transcription? (Or has that already happened?) I had to convert some PDFs and MP3s to text recently and was struck by the vast difference in output quality. Whisper's transcription was near-flawless, all the OCR softwares I tried struggled with formatting, missed words…

You might enjoy this breakdown of the lengths one person went through to take advantage of the iOS vision API and creating a local web service for transcribing some very challenging memes: https://findthatmeme.com/blog/2023/01/08/image-stacks-and-ip... discussed on HN: https://news.ycombinator.com/item?id=34315782

This is so good - thanks for sharing this!

Re: Veo

#246

all of this stuff i'll believe when it's ready for public release 1. safety measures lead to huge quality reductions 2. the devil's in the details. you can make me 1 million videos which look 99% realistic, but it's useless. consumers can pick it instantly, and it's a gigantic turn-off for any brand

There'll always be a market for cheap low-quality videos, and vice versa always a market for shockingly high quality videos. K. Asif's Mughal-e-Azham had enormous ticket sales and a huge budget spending on all sorts of stuff, like actual gold jewelry to make the actors feel that they were important despite the film being black and white.

No matter how good AI gets, it will never be the highest budget. Hell, even technically more accurate quartz watches cannot compete price wise with mechanical masterpiece watches of lower accuracy

Re: Veo

#247
post #182
post #114

The amount of negativity in these comments is astounding. Congrats to the teams at Google on what they have built, and hoping for more competition and progress in this space.

You have to give Google credit as they went against the OpenAI fanatics, Google doomsday crowd and some of the permanent critics (who won't disclose they invested in OpenAI's secondary share sale) that believe that Google can't keep up. In fact, they already did. What OpenAI announced was nothing that Google could not do already. The top comments around Sora vs Veo suggesting that Google was falling behind, given the…

I have no doubt about Google's capabilities in AI, my doubt lies on the productization part. I don't think they can produce something that will not be a complete mess

Re: Veo

#248

Earlier quoted context omitted.

> What OpenAI announced was nothing that Google could not do already I don’t think I’ve seen serious criticism of Google’s abilities. Apple didn’t release anything that Xerox or IBM couldn’t do. The difference is they didn’t. Google’s problem has always been in product follow through. In this case, I fault them for having the sole action item be a buried waitlist request and two new brands (Veo and VideoFX) for one u…

> Google’s problem has always been in product follow through. Google is large enough to not care about small opportunities. It ends up focusing on bigger opportunities that only it can execute well. Google's ability to shut down products that dont work is an insult to user but a very good corporate strategy and they deserve kudos for that. Now, coming back to the "follow through". Google Search, Gmail, Chrome, Androi…

> coming back to the "follow through". Google Search, Gmail, Chrome, Android, Photos, Drive, Cloud etc. all are excellent examples of Google's long term commitment

Do you have any examples of something they launched in the last decade?

Re: Veo

#249
post #85

Earlier quoted context omitted.

OpenAI hardly released gpt-4o. The demo yesterday was clearly a rushed response to I/O. It’s quite possible that Google will ship multi-modality features faster than OpenAI will.

What do you mean? Everyone has access to the gpt-4o model right now through ChatGPT and the API. Sure we don't have voice-to-voice but we have a lot more than what Google has promised.

I don't have access to gpt-4o via ChatGPT
Post reply on HN