Earlier quoted context omitted.
> You could run it on a cluster of nodes Not sure this is a MBP either.
Not even a cluster of Mac Pros could run a dense 5T parameter model with RDMA, to my knowledge.
Microsoft and OpenAI end their exclusive and revenue-sharing deal
791–800 of 915 posts
Re: Microsoft and OpenAI end their exclusive and revenue-sharing deal
#792Re: Microsoft and OpenAI end their exclusive and revenue-sharing deal
#793Earlier quoted context omitted.
We already know that same problem has been examined by many credible mathematicians already and couldn't be solved by any of them yet. Why are we expecting AGI to one shot it? Can't we have an AGI that can fails occasionally to solve some math problem? Is the expectation of AGI to be all knowing? By the way I agree that AGI is not around the corner or I am not arguing any of the llm s are "thinking machines". It's ju…
People want to believe in magic so they will find excuses to do so. Computers have been proving theorems for a long time now but Isabelle/HOL didn't have the marketing budget of OpenAI so people didn't care. Now that Sam Altman is doing the marketing people all of a sudden care about proving theorems.
Re: Microsoft and OpenAI end their exclusive and revenue-sharing deal
#794Earlier quoted context omitted.
One is caused by the other. Amazons engineers decided to split the interface in a “user hostile” manner with the stated purpose of increasing reliability… which didn’t materialise. The clunky UI did. Or maybe you can provide a better explanation for why users had to “hunt” through hundreds(!) of product-region combinations to find that last lingering service they were getting billed $0.01 a month for? This just doesn…
One of the things I find about AWS is that every service UI feels different. It's like every service was designed by a totally different team. For all its flaws at least Azure has consistent UI.
You could argue now that that's no excuse anymore given it's one of the most valuable companies in the world, but that would dismiss the fact they have other priorities than a complete UI overhaul for consistency, and that rewrites are very dangerous, for instance people are already used to the UX pitfalls in the console, it's the devil they know, and changing that will be upsetting to the vast majority of users.
So there you have it. You know what you are getting into, AWS is a behemoth and it's 2026. Don't use the console like it's 2010. Use IaC for any nontrivial work, otherwise you only have yourself to blame.
Re: Microsoft and OpenAI end their exclusive and revenue-sharing deal
#795Earlier quoted context omitted.
People had this "why you probably can't run a GPT-4 (or even GPT-3.5) class model on your MBP anytime soon" conversation before. Today's LLMs are able pack much more capabilities into fewer parameters compared to 2023. We might still be at the very rudimentary phase of this technology there are low-hanging efficiency gains to be had left and right. These models consume many orders of magnitude more energy than a huma…
There is a huge gap between "in two years" and "theoretically possible"
Re: Microsoft and OpenAI end their exclusive and revenue-sharing deal
#796Earlier quoted context omitted.
Companies are starting this year with an agentic layer. We will see how this will affect broader areas
Yeah and every year before there was another poster telling me the next model iteration would be enough.
Than suddenly one model update moves it from 80% to 85% and now 30% of the market wants to use it.
Then it might be already too late to act like using it to your advantage, being a valuable expert or deciding things long term based on the new state of affairs.
Re: Microsoft and OpenAI end their exclusive and revenue-sharing deal
#797Earlier quoted context omitted.
Interesting. Why? My current mental model is that AMD chips are just a bit behind, so, less efficient, but no biggie. Do labs even use CUDA?
This is somewhat out of date (Dec 2024), but gives you some idea of how far behind AMD was then: https://newsletter.semianalysis.com/p/mi300x-vs-h100-vs-h200... Pull quotes: AMD’s software experience is riddled with bugs rendering out of the box training with AMD is impossible. We were hopeful that AMD could emerge as a strong competitor to NVIDIA in training workloads, but, as of today, this is unfortunately not the…
Re: Microsoft and OpenAI end their exclusive and revenue-sharing deal
#798Microsoft and OpenAI quietly killed the AGI clause. The provision that decided what happens when OpenAI builds human-level intelligence, gone. Six months ago that was the most important sentence in tech. Now it's a footnote in a revenu restructuring. Tells you everything about where the AGI conversation actually is.
Re: Microsoft and OpenAI end their exclusive and revenue-sharing deal
#799Earlier quoted context omitted.
Apple is basically in the same boat as AMD and Intel. They have a weak, raster-focused GPU architecture that doesn't scale to 100B+ inference workloads and especially struggles with large context prefill. TPUs smoke them on inference, and Nvidia hardware is far-and-away more efficient for training.
What do TPUs do to improve on GPUs at inference?
Re: Microsoft and OpenAI end their exclusive and revenue-sharing deal
#800Earlier quoted context omitted.
I was once almost fired for saying a little too much in an HN comment about pentesting. Being dragged into an office and given a dressing-down for posting was quite traumatic. The central issue (or so they claimed) was that people might misconstrue my comment as representing the company I was at. So yeah, I don’t understand why people are making fun of this. It’s serious. On the other hand, they were so uptight that…
> On the other hand, they were so uptight that I’m not sure “opinions are my own” would have prevented it. In my experience it didn't matter at all, they considered "you work for us, its known you work for us, therefore your opinions reflect on us". Absolute nonsense, they don't pay me for 24 hours of the day. I told them where they can stick it (politely) and got a new job.