Earlier quoted context omitted.
our only hope for verifying truth in the future is that state officials give their speeches while doing kick flips and frontside 360s.
Maybe they will do more in person talks, I guess. Back to the old times.
Veo 2: Our video generation model
171–180 of 342 posts
Re: Veo 2: Our video generation model
#172I got access to the preview, here's what it gave me for "A pelican riding a bicycle along a coastal path overlooking a harbor" - this video has all four versions shown: https://static.simonwillison.net/static/2024/pelicans-on-bic... Of the four two were a pelican riding a bicycle. One was a pelican just running along the road, one was a pelican perched on a stationary bicycle, and one had the pelican wearing a weird…
There's another important contender in the space: Hunyuan model from Tencent My company (Nim) is hosting Hunyuan model, so here's a quick test (first attempt) at "pelican riding a bycicle" via Hunyuan on Nim: https://nim.video/explore/OGs4EM3MIpW8 I think it's as good, if not better than Sora / Veo
Re: Veo 2: Our video generation model
#173Earlier quoted context omitted.
I think people are just fundamentally not willing to attribute intelligence to things that can't have conversations. This is why the incredible belief was possible that babies or dogs don't feel pain. Once the AI is given some long term memory all of these ideas that AI is just a parrot will suddenly be gone and I personally think that it will probably be pretty easy to give robots memories and their own personal mot…
It is also the corollary: we tend to attribute intelligence to things merely because it can have conversations from the first golden era of AI in 1960's that is always the case. Mimicking more patterns like emotion and motivation may be better user experience, it doesn't make the machine any smarter, just a better mime. Your thesis is that as we mimic reality more and more the differences will not matter, this is a i…
It's beginning to happen. Anthropic hired their first AI welfare researcher from Eleos AI, which is an organization specifically dedicated to investigating this question: https://eleosai.org/
Re: Veo 2: Our video generation model
#174I got access to the preview, here's what it gave me for "A pelican riding a bicycle along a coastal path overlooking a harbor" - this video has all four versions shown: https://static.simonwillison.net/static/2024/pelicans-on-bic... Of the four two were a pelican riding a bicycle. One was a pelican just running along the road, one was a pelican perched on a stationary bicycle, and one had the pelican wearing a weird…
There's another important contender in the space: Hunyuan model from Tencent My company (Nim) is hosting Hunyuan model, so here's a quick test (first attempt) at "pelican riding a bycicle" via Hunyuan on Nim: https://nim.video/explore/OGs4EM3MIpW8 I think it's as good, if not better than Sora / Veo
Turning content blockers off does not make a difference.
Re: Veo 2: Our video generation model
#175Earlier quoted context omitted.
There's another important contender in the space: Hunyuan model from Tencent My company (Nim) is hosting Hunyuan model, so here's a quick test (first attempt) at "pelican riding a bycicle" via Hunyuan on Nim: https://nim.video/explore/OGs4EM3MIpW8 I think it's as good, if not better than Sora / Veo
FYI your website shows me a static image on iOS 18.2 Safari. Strangely, the progress bar still appears to “loop,” but the bird isn’t moving at all. Turning content blockers off does not make a difference.
Re: Veo 2: Our video generation model
#176Superficially impressive but what is the actual use case of the present state of the art? It makes 10-second demos, fine. But can a producer get a second shot of the same scene and the same characters, with visual continuity? Or a third, etc? In other words, can it be used to create a coherent movie --even a 60-second commercial -- with multiple shots having continuity of faces, backgrounds, and lighting? This quote…
Re: Veo 2: Our video generation model
#177Earlier quoted context omitted.
That's the conventional take, but (as far as I can tell), the TPU program was also started under Sundar, which would have been a bold investment at the time, and looks like absolute genius in retrospect. OpenAI might be well-capitalized, but they're (1) bleeding money, (2) no clear path to profitability, and (3) competing head-to-head with a behemoth who can profitably provide a similar offering at 10-20x cheaper (li…
No they’ve been acquiring layers upon layers of middle management for a decade. That’s the core issue, and they’ve also pissed off a non-zero percentage of top talent by ditching what still existed of Google culture and going full “Corporate Megacorp” a few years ago. Google is having to pay a ton to retain the talent they have left and it’s often not enough.
Google's biggest threat isn't OpenAI. It's the FTC (which I admit is a very real danger).
* from a developer/platform perspective, at least. The "consumer" facing side of things (e.g. the AI Studio UI) is still pretty awful.
Re: Veo 2: Our video generation model
#178Earlier quoted context omitted.
No they’ve been acquiring layers upon layers of middle management for a decade. That’s the core issue, and they’ve also pissed off a non-zero percentage of top talent by ditching what still existed of Google culture and going full “Corporate Megacorp” a few years ago. Google is having to pay a ton to retain the talent they have left and it’s often not enough.
That might be true, but I'm not sure it will matter that much. They've put out two very compelling* products in the last month (Flash 2 & Veo), and hints that another Gemini Pro model is in the pipeline. When it comes to AI, they're in a very good position, no matter what middle-management shenanigans are going on behind the scenes. Their core ad business is also so absurdly profitable that overpaying for talent won'…
Re: Veo 2: Our video generation model
#179just to remind everyone that state of the art was Will Smith Eating Spaghetti in April of 2023 https://arstechnica.com/information-technology/2023/03/yes-v... We're not even done with 2024. Just imagine what's waiting for us in 2025.
But it's the same thing just at a higher fidelity. Which is impressive don't get me wrong. But they are also kinda bad looking. Like even there good examples have so many issues. I just don't see how this gets extrapolated into the ideas in various posts like full length movies, custom TV shows and holodecks or whatever else people dream up. Do we have any examples of tech that just kept improving at exponential or l…
SD Cards?
Re: Veo 2: Our video generation model
#180I got access to the preview, here's what it gave me for "A pelican riding a bicycle along a coastal path overlooking a harbor" - this video has all four versions shown: https://static.simonwillison.net/static/2024/pelicans-on-bic... Of the four two were a pelican riding a bicycle. One was a pelican just running along the road, one was a pelican perched on a stationary bicycle, and one had the pelican wearing a weird…
It looks much better than Sora but still kind of in uncanny valley