OpenAI Jalapeño: Better than Nvidia Blackwell
181–190 of 389 posts
Re: OpenAI Jalapeño: Better than Nvidia Blackwell
#182I think they talked about this being general purpose chip but I would think that Anthropic/OpenAI are at the scale now they could bake LLM weights into chips themselves. For example, GPT Sol baked into a custom chip run for $100M that runs 10x as fast and 10x as cheap should pay for itself as long as the chip is useful for long enough. While 2 years ago nothing was useful more than 1 year long, there are many older m…
Taalas needed a giant chip (6nm) for an 8B model.
At best you could use a more advanced node to try to put a MoE model across several chips working together, but you can’t have GPT Sol size models on a single chip like that.
Re: OpenAI Jalapeño: Better than Nvidia Blackwell
#183Re: OpenAI Jalapeño: Better than Nvidia Blackwell
#184I love how now you have to consider the possible s** posting motivation behind analysis of a trillion dollar industry being conducted at a world-class level by a bunch of ex Reddit and 4Chan adjacent mods -- it's one of the best stories in AI that SemiAnalysis is not cut from the same cloth as Gartner McKinsey et al
lol. lmao even.
Have you seen the quality of their output? I'd take Claude or ChatGPT Free Tier over advice from McKinsey these days.
Re: OpenAI Jalapeño: Better than Nvidia Blackwell
#185Re: OpenAI Jalapeño: Better than Nvidia Blackwell
#186I hadn't seen the token/Joules comparison with human speech before. Humans are still 22x more efficient, which is not that far considering the rate of progress in this area.
Based on a human output rate of 3.3 tok/s, which seems questionable as a means of comparison
Re: OpenAI Jalapeño: Better than Nvidia Blackwell
#187I hadn't seen the token/Joules comparison with human speech before. Humans are still 22x more efficient, which is not that far considering the rate of progress in this area.
I wonder how that stacks up if you consider all the time you have to keep the body alive when it’s not actively producing “tokens”.
Productivity is not the only reason to let these meatbags burn oxygen.
Re: OpenAI Jalapeño: Better than Nvidia Blackwell
#188Earlier quoted context omitted.
We should be mindful of the context that many of these providers VERY likely have been selling their subscriptions at a substantial loss So as much as i agree “more profits to stakeholders screw the customer”, i think its more of an emergency to get to profitability before the music stops.
> We should be mindful of the context that many of these providers VERY likely have been selling their subscriptions at a substantial loss. what makes you think this?
Re: OpenAI Jalapeño: Better than Nvidia Blackwell
#189Earlier quoted context omitted.
I can totally see how ternary would work from a physical implementation perspective but I have a really hard time visualizing anything using base-e, can you explain how such a thing would work in practice?
No it's impossible. But it would be optimal!
Re: OpenAI Jalapeño: Better than Nvidia Blackwell
#190I think they talked about this being general purpose chip but I would think that Anthropic/OpenAI are at the scale now they could bake LLM weights into chips themselves. For example, GPT Sol baked into a custom chip run for $100M that runs 10x as fast and 10x as cheap should pay for itself as long as the chip is useful for long enough. While 2 years ago nothing was useful more than 1 year long, there are many older m…