Live data from Hacker News

Gemini Omni 1.1 Flash

blog.google

221–230 of 252 posts

Re: Gemini Omni 1.1 Flash

#223

I let myself get mildly excited with the last Omni release, but it turns out it (and this one) can't do the one practical thing I want - Sync generated video to provided pre-existing audio. Meanwhile, I'm happily using Minimax H3 locally on my 12Gb 4070RTX to finally finish the lip syncing to recorded dialog on my abandoned 20 year old Flash animation hobby projects.

> Minimax H3 locally on my 12Gb 4070RTX Minimax H3 is about 240Gb alone, how do you do? How much quantised is it, and how good are the results?

Not op, but where are you getting that number from? Even the full 16-bit precision model is only about 66GB.

Most people are running the stock release of Minimax H3 using the INT8 quant and it's about ~20GB.

https://huggingface.co/Comfy-Org/MiniMax-H3/tree/main/diffus...

https://docs.comfy.org/tutorials/video/minimax/minimax-h3

Re: Gemini Omni 1.1 Flash

#226
post #23
post #4

Earlier quoted context omitted.

Many screen and voice actors are unionized, and the unions have been striking and bargaining specifcially over these points. For example: https://sites.suffolk.edu/jhtl/2025/10/30/game-over-for-unau... Software developers, unfortunately, have been convinced that they don't need to unionize, so have no collective bargaining power for dealing with situations like this.

Personally, I just can’t find it in me to be that self-interested. It’s how other people oppose buildings near them because it blocks their view. I really don’t want to stop other people from writing software. If they want to use AI to do it so be it. That’ll be somewhat detrimental to me perhaps (which isn’t certain), but that’s okay. I’d rather attempt to adapt to a changing world.

Right, but in your apathy there is a well organised group of capitalists who are vying to instrumentalise your existence. A union is just levelling the playing field.

Re: Gemini Omni 1.1 Flash

#227
post #24

Interesting that OpenAI abandoned Sora entirely but Google are continuing to invest heavily in their own video generation. Maybe because they see video generation as key to developing "world models"?

Google has always been committed to multimodal. And, Veo and omni simply were better than Sora And, Youtube is huge both as a place where video contents goes and where can be trained from. Microdramas are starting to become a real category--14 Billion USD, 90% of it made with AI. Chinese video models can be more immediately impressive, but none of them come close to beat the value of Google's Flow. Especially when yo…

> Microdramas are starting to become a real category--14 Billion USD

Almost all concrete english-language info I can find about microdramas is astroturfed to hell by consultants and "independent" industry publications. Wikipedia's citations for 2025 revenue are 'Reel Reel' and 'Duanju News'. Neither cite their source, though they are likely just regurgitating predictions from Omdia, a media consultancy.

> Real Reel™ works with companies across entertainment, technology and the creator economy to build relevant industry conversations around mobile-first storytelling.

Duanju News is published by 'Studio Phocéen', who run their own microdrama production house.

The Omdia revenue estimates, which is where the projected $14 billion comes from, are completely unsubstantiated as far as I can tell. They even go so far as to make "according to new research" a link that when followed sends you to their generic "/advance-your-business/media-and-entertainment" sales pitch.

https://omdia.tech.informa.com/pr/2025/oct/microdramas-to-ge...

I don't have any special insights here, it all just smells a bit like McKinsey's "The metaverse will be worth $5 trillion by 2030, you better not miss out!!! Hire our 22 year old slide deck experts today"

Re: Gemini Omni 1.1 Flash

#228
Does the last video demo give a totalitarian AI apocalypse feeling to anyone else?

"Omni" -> everwhere/all

"The final chapter will end with a choice. The choice between holding on to the past, or letting go. What will you choose?"

Re: Gemini Omni 1.1 Flash

#230

Earlier quoted context omitted.

AI / LLM is about more than agentic coding. It is one of the least interesting use cases to me, thinking more broadly. HN may be over-indexed on it.

I would agree with you on AI/LLM being more than agentic coding but at the same time, I think there's more nuance. For example, PDF's and powerpoints can be generated using agentic coding by things like https://bento.page or other ways of generating them in an agentic coding fashion. A lot of browser automation could/is also done by agentic coding. It can also help them set up and configure self hosted software with…

I get that. I use agents a lot and LLMs often reason with code. It is valuable. I just think the floor is a lot lower for general reasoning and common tasks like that. And in 6-12 months it won’t matter. Google will publish better models. The temporal distortion of how long a Sol or a Fable has existed is real. No one is suddenly missing out on some giant competitive edge because their model is a few months behind. I feel like it’s all just going to normalize and things other than how well your model can write code will matter more and more in 12 to 24 months.
Post reply on HN