Live data from Hacker News

H3-metal – Native MiniMax-H3 inference for Apple Silicon

github.com

41–50 of 108 posts

Re: H3-metal – Native MiniMax-H3 inference for Apple Silicon

#41

Alright I’ve been afraid to ask but have been having trouble finding What are some adult entertainment workflows in comfyui, I need best loras, best prompts to start with and the communities, are they on telegram or something?

This is a healthy question. We want to use these tools for regular old human needs and desires.

Re: H3-metal – Native MiniMax-H3 inference for Apple Silicon

#42
post #29

Earlier quoted context omitted.

Being a world class talent is independent of financial situation.

Of course it isn’t. If you can’t afford to eat you can’t achieve any potential you might have. Financial stability is a gamechanger for everyone.

That's stupid, if you're truly talented you'll solve the financial stuff in order to pursue whatever you want to do - if you don't then that's on you.

Re: H3-metal – Native MiniMax-H3 inference for Apple Silicon

#43

On my 128GB M4 Max Mac Studio, generating a 15s 480p video with MiniMax H3 in ComfyUI takes an hour and a half. Put Codex to work on deploying it now, hoping the speed can improve quite a lot :-) Thanks anyway

Please keep up posted about the results!

First batch of quick test results: approximately 1/5 speed improvement

Re: H3-metal – Native MiniMax-H3 inference for Apple Silicon

#45

On my 128GB M4 Max Mac Studio, generating a 15s 480p video with MiniMax H3 in ComfyUI takes an hour and a half. Put Codex to work on deploying it now, hoping the speed can improve quite a lot :-) Thanks anyway

> On my 128GB M4 Max Mac Studio, generating a 15s 480p video with MiniMax H3 in ComfyUI takes an hour and a half.

That's crazy, a RTX Pro 6000 does that in in 2-3 minutes (give or take, depending on your exact settings). LLMs don't make the difference between standalone GPU vs unified memory + CPU so obvious as diffusion models seems to do.

Re: H3-metal – Native MiniMax-H3 inference for Apple Silicon

#46
post #9

This is where the DGX spark makes up a bit of the ground it loses on llm work, diffusion and cuda go together like peanut butter and jelly.

cough DiffusionGemma cough

Seriously, very dumb model compared to what you can run locally, but holy moly is it FAST on one GPU, seriously impressive. Can't wait for those to be scaled up a bit to fit perfectly within 96GB VRAM, then they'll be competitive.

Re: H3-metal – Native MiniMax-H3 inference for Apple Silicon

#47

This still requires 128Gb of memory, right? Me and my lowly 96Gb, like a commoner; missing out on the fun.

you should have a look at https://github.com/deepbeepmeep/Wan2GP which is the goto tool for "gpu poor", although as people below already pointed out you should be fine with comfyui's standard setup aswell

Re: H3-metal – Native MiniMax-H3 inference for Apple Silicon

#48

This still requires 128Gb of memory, right? Me and my lowly 96Gb, like a commoner; missing out on the fun.

It shouldn't? Unless you're using BF16 for all weights (I'm using NVFP4 for the text encoder, otherwise everything BF16 (and audio F32)) you'll fit it all within 96GB VRAM, bugs non-with-standing :) I've been fitting this within 96GB VRAM without issues.

Re: H3-metal – Native MiniMax-H3 inference for Apple Silicon

#49

This still requires 128Gb of memory, right? Me and my lowly 96Gb, like a commoner; missing out on the fun.

you should have a look at https://github.com/deepbeepmeep/Wan2GP which is the goto tool for "gpu poor", although as people below already pointed out you should be fine with comfyui's standard setup aswell

First, I think they're not even talking about GPUs, this is macOS hardware so unified memory. Secondly, if they were talking about GPUs, then 96GB VRAM is hardly what people refer to when they say "gpu poor".

Re: H3-metal – Native MiniMax-H3 inference for Apple Silicon

#50

Alright I’ve been afraid to ask but have been having trouble finding What are some adult entertainment workflows in comfyui, I need best loras, best prompts to start with and the communities, are they on telegram or something?

a friend told me there's a reddit called: unstable diffusion
Post reply on HN