Groq runs Mixtral 8x7B-32k with 500 T/s
21–30 of 482 posts
Re: Groq runs Mixtral 8x7B-32k with 500 T/s
#22Hi folks, I work for Groq. Feel free to ask me any questions. (If you check my HN post history you'll see I post a lot about Haskell. That's right, part of Groq's compilation pipeline is written in Haskell!)
Re: Groq runs Mixtral 8x7B-32k with 500 T/s
#23Re: Groq runs Mixtral 8x7B-32k with 500 T/s
#24Hi folks, I work for Groq. Feel free to ask me any questions. (If you check my HN post history you'll see I post a lot about Haskell. That's right, part of Groq's compilation pipeline is written in Haskell!)
Re: Groq runs Mixtral 8x7B-32k with 500 T/s
#25Earlier quoted context omitted.
They’re for sale on Mouser for $20625 each https://www.mouser.com/ProductDetail/BittWare/RS-GQ-GC1-0109... At that price 568 chips would be $11.7M
That seems to be per card instead of chip. I would expect it has multiple chips on a single card.
> Accelerator Cards GroqCard low latency AI/ML Inference PCIe accelerator card with single GroqChip
Re: Groq runs Mixtral 8x7B-32k with 500 T/s
#26Hi folks, I work for Groq. Feel free to ask me any questions. (If you check my HN post history you'll see I post a lot about Haskell. That's right, part of Groq's compilation pipeline is written in Haskell!)
@tome for the deterministic system, what if the timing for one chip/part is off due to manufacturing/environmental factors (e.g., temperature) ? How does the system handle this?
Re: Groq runs Mixtral 8x7B-32k with 500 T/s
#27Earlier quoted context omitted.
That seems to be per card instead of chip. I would expect it has multiple chips on a single card.
From the description that doesn't seem to be the case, but I don't know this product well > Accelerator Cards GroqCard low latency AI/ML Inference PCIe accelerator card with single GroqChip
Re: Groq runs Mixtral 8x7B-32k with 500 T/s
#28Earlier quoted context omitted.
Yeah, it's nothing to do with Elon and we (Groq) had the name first. It's a natural choice of name for something in the field of AI because of the connections to the hacker ethos, but we have the trademark and Elon doesn't. https://wow.groq.com/hey-elon-its-time-to-cease-de-grok/
Can't Chamath (he's one of your investors, right), do a thing there? Every person I pitch Groq to is confused and thinks its about Elons unspectacular LLM.
Re: Groq runs Mixtral 8x7B-32k with 500 T/s
#29Re: Groq runs Mixtral 8x7B-32k with 500 T/s
#30Hi folks, I work for Groq. Feel free to ask me any questions. (If you check my HN post history you'll see I post a lot about Haskell. That's right, part of Groq's compilation pipeline is written in Haskell!)
@tome for the deterministic system, what if the timing for one chip/part is off due to manufacturing/environmental factors (e.g., temperature) ? How does the system handle this?