Live data from Hacker News

H3-metal – Native MiniMax-H3 inference for Apple Silicon

github.com

71–80 of 108 posts

Re: H3-metal – Native MiniMax-H3 inference for Apple Silicon

#71

I've been using MiniMax H3 on my M5 Pro 64GB MacBook Pro through ComfyUI. It works extremely well. I had to modify the default ComfyUI workflows to use a GGUF quant (city96's ComfyUI-GGUF custom node, UnetLoaderGGUF in place of the stock loader) [0]. I use the model labeled Q5_K_M. There is Q8_0 available as well, which is 34GB and fits fine in 64GB unified memory if you keep resolution modest. The main issue is spee…

GGUF is outdated in the latest versions of Comfy-UI. If you want a good balance of size, speed and quality you should use the int8_convrot model from the official Comfy Org Repo https://huggingface.co/Comfy-Org/MiniMax-H3/tree/main/diffus...

Re: H3-metal – Native MiniMax-H3 inference for Apple Silicon

#72

wow antirez does not sleep

when you have enough money to not have to worry about anything, you can go back to your hobbies. in this case, his hobby is programming.

Redis, hping and dump1090 were all side projects he started/written while having a full time job.

Your comment really sounds like "many other people would do A and B if they just had time and money to do so", but he's been doing so since time and money were major constraints.

Re: H3-metal – Native MiniMax-H3 inference for Apple Silicon

#73

I've been using MiniMax H3 on my M5 Pro 64GB MacBook Pro through ComfyUI. It works extremely well. I had to modify the default ComfyUI workflows to use a GGUF quant (city96's ComfyUI-GGUF custom node, UnetLoaderGGUF in place of the stock loader) [0]. I use the model labeled Q5_K_M. There is Q8_0 available as well, which is 34GB and fits fine in 64GB unified memory if you keep resolution modest. The main issue is spee…

[deleted]

Re: H3-metal – Native MiniMax-H3 inference for Apple Silicon

#74

On my 128GB M4 Max Mac Studio, generating a 15s 480p video with MiniMax H3 in ComfyUI takes an hour and a half. Put Codex to work on deploying it now, hoping the speed can improve quite a lot :-) Thanks anyway

> On my 128GB M4 Max Mac Studio, generating a 15s 480p video with MiniMax H3 in ComfyUI takes an hour and a half. That's crazy, a RTX Pro 6000 does that in in 2-3 minutes (give or take, depending on your exact settings). LLMs don't make the difference between standalone GPU vs unified memory + CPU so obvious as diffusion models seems to do.

LLM prompt processing and diffusion models are compute bound, while LLM token generation is memory bandwidth bound.

Re: H3-metal – Native MiniMax-H3 inference for Apple Silicon

#75

I've been using MiniMax H3 on my M5 Pro 64GB MacBook Pro through ComfyUI. It works extremely well. I had to modify the default ComfyUI workflows to use a GGUF quant (city96's ComfyUI-GGUF custom node, UnetLoaderGGUF in place of the stock loader) [0]. I use the model labeled Q5_K_M. There is Q8_0 available as well, which is 34GB and fits fine in 64GB unified memory if you keep resolution modest. The main issue is spee…

What is the quality of the output like compared to something like Veo?

It's better than all private video models like Veo. Yes, I too am incredulous they released this open weights. It's a VERY disruptive model in all the good ways.

Re: H3-metal – Native MiniMax-H3 inference for Apple Silicon

#76
post #4

How similar are Jeff Dean and Salvatore Sanfilippo?

My favorite Jeff Dean fact is that he’s also antirez. Which reminds me of my favorite Salvatore Sanfilippo fact. He’s also Jeff Dean

Identify theft is not a joke Jim!

Re: H3-metal – Native MiniMax-H3 inference for Apple Silicon

#77
post #39

Earlier quoted context omitted.

There will Reddit subs for it though couldn’t tell you which off top of my head I’d personally steer clear of messaging platforms for this - who knows what one might stumble into there

> I’d personally steer clear of messaging platforms for this - who knows what one might stumble into there Personally I have no interest, but sometime browse stuff out of curiosity. But this got more of my curiosity, what kind of "stuff" are you implying they might stumble upon on the open, public internet? Sure, some NSFW, horror and otherwise weird stuff is there, especially around AI generation, but hardly somethi…

One thing I read on this topic on reddit is never EVER use the word "girl" when prompting H3. So CSAM probably.

Re: H3-metal – Native MiniMax-H3 inference for Apple Silicon

#78
I'm looking to setup a way to create images for my own instagram marketing. I do not care how long it takes to make 10 variations of a post as that speed would still be faster than me making it.

Does this model work with ComfyUI easily? Can I just download it?

Post reply on HN