The path to ubiquitous AI (17k tokens/sec)
1–10 of 471 posts
Re: The path to ubiquitous AI (17k tokens/sec)
#2Re: The path to ubiquitous AI (17k tokens/sec)
#3Re: The path to ubiquitous AI (17k tokens/sec)
#4Re: The path to ubiquitous AI (17k tokens/sec)
#5I'll take one with a frontier model please, for my local coding and home ai needs..
Re: The path to ubiquitous AI (17k tokens/sec)
#6Re: The path to ubiquitous AI (17k tokens/sec)
#7Show me something at a model size 80GB+ or this feels like "positive results in mice"
Re: The path to ubiquitous AI (17k tokens/sec)
#8This requires 10 chips for an 8 billion q3 param model. 2.4kW.
10 reticle sized chips on TSMC N6. Basically 10x Nvidia H100 GPUs.
Model is etched onto the silicon chip. So can’t change anything about the model after the chip has been designed and manufactured.
Interesting design for niche applications.
What is a task that is extremely high value, only require a small model intelligence, require tremendous speed, is ok to run on a cloud due to power requirements, AND will be used for years without change since the model is etched into silicon?
Re: The path to ubiquitous AI (17k tokens/sec)
#9So I'm guessing this is some kind of weights as ROM type of thing? At least that's how I interpret the product page, or maybe even a sort of ROM type thing that you can only access by doing matrix multiplies.