I've been using MiniMax H3 on my M5 Pro 64GB MacBook Pro through ComfyUI. It works extremely well. I had to modify the default ComfyUI workflows to use a GGUF quant (city96's ComfyUI-GGUF custom node, UnetLoaderGGUF in place of the stock loader) [0]. I use the model labeled Q5_K_M. There is Q8_0 available as well, which is 34GB and fits fine in 64GB unified memory if you keep resolution modest. The main issue is spee…
H3-metal – Native MiniMax-H3 inference for Apple Silicon
71–80 of 108 posts
Re: H3-metal – Native MiniMax-H3 inference for Apple Silicon
#72wow antirez does not sleep
when you have enough money to not have to worry about anything, you can go back to your hobbies. in this case, his hobby is programming.
Your comment really sounds like "many other people would do A and B if they just had time and money to do so", but he's been doing so since time and money were major constraints.
Re: H3-metal – Native MiniMax-H3 inference for Apple Silicon
#73I've been using MiniMax H3 on my M5 Pro 64GB MacBook Pro through ComfyUI. It works extremely well. I had to modify the default ComfyUI workflows to use a GGUF quant (city96's ComfyUI-GGUF custom node, UnetLoaderGGUF in place of the stock loader) [0]. I use the model labeled Q5_K_M. There is Q8_0 available as well, which is 34GB and fits fine in 64GB unified memory if you keep resolution modest. The main issue is spee…
Re: H3-metal – Native MiniMax-H3 inference for Apple Silicon
#74On my 128GB M4 Max Mac Studio, generating a 15s 480p video with MiniMax H3 in ComfyUI takes an hour and a half. Put Codex to work on deploying it now, hoping the speed can improve quite a lot :-) Thanks anyway
> On my 128GB M4 Max Mac Studio, generating a 15s 480p video with MiniMax H3 in ComfyUI takes an hour and a half. That's crazy, a RTX Pro 6000 does that in in 2-3 minutes (give or take, depending on your exact settings). LLMs don't make the difference between standalone GPU vs unified memory + CPU so obvious as diffusion models seems to do.
Re: H3-metal – Native MiniMax-H3 inference for Apple Silicon
#75I've been using MiniMax H3 on my M5 Pro 64GB MacBook Pro through ComfyUI. It works extremely well. I had to modify the default ComfyUI workflows to use a GGUF quant (city96's ComfyUI-GGUF custom node, UnetLoaderGGUF in place of the stock loader) [0]. I use the model labeled Q5_K_M. There is Q8_0 available as well, which is 34GB and fits fine in 64GB unified memory if you keep resolution modest. The main issue is spee…
What is the quality of the output like compared to something like Veo?
Re: H3-metal – Native MiniMax-H3 inference for Apple Silicon
#76Re: H3-metal – Native MiniMax-H3 inference for Apple Silicon
#77Earlier quoted context omitted.
There will Reddit subs for it though couldn’t tell you which off top of my head I’d personally steer clear of messaging platforms for this - who knows what one might stumble into there
> I’d personally steer clear of messaging platforms for this - who knows what one might stumble into there Personally I have no interest, but sometime browse stuff out of curiosity. But this got more of my curiosity, what kind of "stuff" are you implying they might stumble upon on the open, public internet? Sure, some NSFW, horror and otherwise weird stuff is there, especially around AI generation, but hardly somethi…
Re: H3-metal – Native MiniMax-H3 inference for Apple Silicon
#78Does this model work with ComfyUI easily? Can I just download it?