Live data from Hacker News

MiniMax H3 Day-0 Support in ComfyUI: Open Weights, Native Audio, and 2K Video

blog.comfy.org

11–20 of 100 posts

Re: MiniMax H3 Day-0 Support in ComfyUI: Open Weights, Native Audio, and 2K Video

#13
post #8

Im running this on my 4070ti super (16 gb vram), and it takes 10 minutes for a 10-seconds 480p video. but the results are spectacular.

If you wouldn't mind sharing, what's your Comfy workflow for this? I have the same video card setup and would like to give it a shot.

Re: MiniMax H3 Day-0 Support in ComfyUI: Open Weights, Native Audio, and 2K Video

#14
post #8

Im running this on my 4070ti super (16 gb vram), and it takes 10 minutes for a 10-seconds 480p video. but the results are spectacular.

FWIW on a 5080 16GB it takes 3 minutes for 10 seconds 480p video (the mouse video workflow with length changed from 5 seconds to 10 seconds)

Re: MiniMax H3 Day-0 Support in ComfyUI: Open Weights, Native Audio, and 2K Video

#15
> We found that the model's modulation weights (~40% of the total parameters) could be pruned and replaced with a functionally equivalent lookup table, dramatically shrinking the memory footprint with no loss in output quality.

Is this a common approach to reducing weights with "no loss in output quality", assuming this is true? Seems almost too simple to work. If this is doable, would this be applicable to LLMs as well?

Neat with native frame-to-frame generation, but wonder how easy it is to "link" together clips at the intersection, typically the models kind of lose the "momentum" across these stiches, being able to merge things with frame-to-frame between clips might help with this it feels like.

Re: MiniMax H3 Day-0 Support in ComfyUI: Open Weights, Native Audio, and 2K Video

#16
post #8

Im running this on my 4070ti super (16 gb vram), and it takes 10 minutes for a 10-seconds 480p video. but the results are spectacular.

If you wouldn't mind sharing, what's your Comfy workflow for this? I have the same video card setup and would like to give it a shot.

just the default one in the link for image-to-video

Re: MiniMax H3 Day-0 Support in ComfyUI: Open Weights, Native Audio, and 2K Video

#17
post #10

I saw the samples people have posted. Immediately deleted LTX2 and WAN folders. Those are completely worthless now. There is some debate on the license for those in the US, UK, EU, plus… no comment other than whew those samples though!

"Regions such as the EU, UK, South Korea, and the US are currently developing or enforcing AI-related regulations that may have specific implications for generative video models" You just have to pinkie promise you won't make disney mad and they will send you a licence https://huggingface.co/MiniMaxAI/MiniMax-H3/blob/main/docs/Q...

I do work animations for fun and internal use only (mostly jokes). I MAY reach out to them.

Re: MiniMax H3 Day-0 Support in ComfyUI: Open Weights, Native Audio, and 2K Video

#18
post #8

Im running this on my 4070ti super (16 gb vram), and it takes 10 minutes for a 10-seconds 480p video. but the results are spectacular.

I am particularly curious how multimodal models will work with types of knowledge that are inherently non-text. For example, SOTA LLMs really suck at electronics, especially analog electronics.

Is MiniMax H3 capable of logical / technical reasoning, or is it purely art oriented?

Re: MiniMax H3 Day-0 Support in ComfyUI: Open Weights, Native Audio, and 2K Video

#19
I've said it before and I'll say it again, human directors are still valuable, as they use AI video editing tools to generate the shots they want and put them together in a cohesive way. Previously they might've used film and actors but if they can just prompt the AI (or create workflows as seen with ComfyUI) then they arrange them together just like how an EDM producer doesn't actually play the instruments but instead the creativity is in the arrangement.

I suspect it'll be quite a while until AI gets a good enough aesthetic sense to do this, as even with static HTML websites humans can easily see that it's AI slop.

Post reply on HN