pitch basically boils down to 'just change one line and it works' which sounds too good to be true, but if they actually pull it off at 100k-chip scale, that's genuinely a big deal
The pitch is "just change one line and it works". It is not "just change one line and you will get peak performance on the TPU".
From the text of the blog post:
"Portability doesn't eliminate hardware realities, so TorchTPU facilitates a tiered workflow: establish correct execution first, then use our upcoming deep-dive guidelines to identify and refactor suboptimal architectures, or to inject custom kernels, for optimal hardware utilization."
Is it just me, or does it feel like everyone now uses AI to write any kind of blog? These parts here somehow trigger me: - Enter TorchTPU. As an engineering team, our mandate was to build a stack that leads with usability, portability, and excellent performance. - Engineering the TorchTPU Stack: The Technical Reality - Eager First: Flexibility Without Compromise - The breakthrough, however, is our fused eager mode. -…
I think HN has moved from outright, unabashed llm hate, to a hyperactive cynicism so you can be the HN guy who caught out llm malfeasance. Give it three more years to cook. And I'll go a step further. This audience who claims to hate the unreal/false llm screed is the same set of people who caused tech news to go from Siracusa reviewing macos and Anand reviewing gpu's to car reviews on ars and anandtech being dead. If a tenth of the people who claim to hate shit llm articles were true, Ars and Anandtech wouldn't be dead