OpenAI Jalapeño: Better than Nvidia Blackwell
321–330 of 390 posts
Re: OpenAI Jalapeño: Better than Nvidia Blackwell
#322Re: OpenAI Jalapeño: Better than Nvidia Blackwell
#323Yeah, stopped reading there, this is obviously some deranged sam altman paid blog post, I can't wait for the bubble to pop just so his newly launched chip falls flat on his face.
Re: OpenAI Jalapeño: Better than Nvidia Blackwell
#324Earlier quoted context omitted.
these GPUs make computation faster, I understand as of now maybe all the computation is used to generate yet another junk LinkedIn post or unnecessary RFC, but at some point this craze should settle and we will be left with powerful computation machines, which can be used for computing more useful things
The GPUs being paid for w/ billions in investment will be obsolete and e-waste in a few short years same as a Cray-2 was just a decade after its release. It's fine if you're one of the people selling shovels to gold miners for a while, but sucks to be building houses in the boom town?
In terms of GPUs whole world with 8B people have only couple of viable options: Nvidia, AMD, Intel - and largest part of their inventory is going to enterprises to run those LLMs, and its impacting every consumer / hobby projects, like cheap phones, DIY electronics projects and so on.
I want to have more alternatives on the market
Re: OpenAI Jalapeño: Better than Nvidia Blackwell
#325This semi-analysis article reads a lot more like an OpenAI press release than a real analysis. And to be honest some of the statements seem like just straight up lies - they initially claim they were invited to benchmark it, and then half way down switch to claiming that OpenAI provided all the numbers. This really kind of sucks, because I want to read actual detailed nuanced and credible analysis of what's happening…
Re: OpenAI Jalapeño: Better than Nvidia Blackwell
#326Earlier quoted context omitted.
Imagine Anthropic gives you Opus of 6 months ago but at much higher speeds and much lower cost (that they might or might not pass on). Would you use it?
Yes, this would basically obviate Sonnet and Haiku. If you consider them 1 and 2 generations behind, respectively (that's not really what they are), you can still get a ton out of those older chips. Not to mention people still use older Opus versions happily. (In part because they don't like the new Opus but still, the cost effectiveness is a huge boon.)
I don't know why anyone would use Haiku.
Re: OpenAI Jalapeño: Better than Nvidia Blackwell
#327Earlier quoted context omitted.
The link you cited is not evaporative cooling and a pot of boiling water with a sealed lid on it is a pressure vessel which eventually explodes.
If you were to remove the heat at a sufficient rate by, say, turning the lid into a heat exchanger, you would have a stable system.
I'm not a datacenter engineer, but I used to work in the ski industry. Snowmaking systems use vast quantities of compressed air. It works better if that air is cool. Blowing hot compressed air out of a snow cannon means the air temperature (wet bulb to be specific) needs to be colder to make snow.
Anyways, most air compression stations use water to cool the air, and then evaporative coolers to cool the water. The water is reused, but a ton (not sure of the percentage) is lost into the air. It's more or less a tower with a big fan on top, and water percolates down from the top, being cooled by the air as it goes. The water is then collected and pumped through the system again (but of course has to be always topped up to counteract what was lost to evaporation).
Anyways, long story short is it's most cost effective to just spray water into the air to cool water, as long as water is free/cheap.
Re: OpenAI Jalapeño: Better than Nvidia Blackwell
#328Continued hardware improvements really make it hard for me to believe token prices will not continue to plummet.
Re: OpenAI Jalapeño: Better than Nvidia Blackwell
#329Re: OpenAI Jalapeño: Better than Nvidia Blackwell
#330The article is a bit naive: > However, as previously mentioned, Jalapeño’s results are obtained without speculative decoding and Vera Rubin’s results use speculative decoding. Speculative decoding leads to a ~3-5x reduction in cost per token. When speculative decoding is implemented on Jalapeño, this will enable Jalapeño to serve tokens even more cost effectively. How much speculative decoding improves throughput is…