Earlier quoted context omitted.
Agreed 100%. This guy thinks there's a limit on the demand for intelligence. You think that Fable 7 which can run a billion dollar corporation on its own has no consumer demand just because we have fable 5 at 9k tok/s? Who do you think will be the biggest customer of such a model? Fable 7, obviously.
Sorry, which billion-dollar corporation is Fable running "on its own"?
Kimi K3, Qwen 3.8, and Anthropic's (Potential) Unravelling
311–320 of 349 posts
Re: Kimi K3, Qwen 3.8, and Anthropic's (Potential) Unravelling
#312It's amazing how quickly Fable went from 'Game-changing model that needs to be banned' to 'Yeah it's alright, but OpenAI is also just as good and there are a couple of good open weight alternatives that are equivalent for almost everything' The hype cycles are shortening, perhaps we really are reaching some kind of plateau this time (famous last words)
Re: Kimi K3, Qwen 3.8, and Anthropic's (Potential) Unravelling
#313Upper bound of AI progress - recursive self improvement. In this case AI will be responsible for building better models, making people who own datacenters the winners. Anthropic/OAI is cooked. Lower bound of AI progress - plateau. Progess is slowing, focus is on serving a meaningful peak capability at the lowest possible price. There's been news today that Google is building a Gemini chip with weights baked into sili…
That feels like a very generous framing. There's very little opportunity between the extremes that would paint a convincing outlook of survival for either company.
Perhaps if they were able to scale down their spending drastically they could survive, but that requires acknowledging their current valuations are BS. Doing so is a major risk, that will piss off all share holders. There's also the employees they would need to fire or reduce salary. The shift of focus internally to sustainability would be a major challenge.
Re: Kimi K3, Qwen 3.8, and Anthropic's (Potential) Unravelling
#314The open weight, open architecture releases of the past several days has me more convinced that ultimately, the winner will be whoever burns their models to ASICs fastest. The LLMs themselves are capable of doing some aspects of chip design as evinced by the K3 press release. Furthermore, the frontier models are "good enough" for a wide swathe of tasks and will soon hit that threshold for a good amount of software en…
> the winner will be whoever burns their models to ASICs fastest. It'll obviously be China, and they won't need the bleeding edge of lithography tech to make it happen. Every Chinese smartphone will have something like Sonnet 5, along your car (well, not those of us in the US, but we'll look longingly at pictures of them while we drive whatever the government decides we're allowed to drive in Fortress America). Give…
China won in cloud services? Nope, not even close. They had to clone AWS just to try to keep up.
China won in mobile? Nope. Although they're very competitive there.
China won in search? Nope. Baidu who?
China won in ecommerce? Nope. Their dominance is overwhelmingly domestic.
China won in software? Nope. Windows is US. Android is US. iOS is US. MacOS is US. Linux is US/Europe/Global. Look at the top 50 largest software companies.
China won in silicon? Nope. Look at the top 50 companies.
China won in .... India and Latin America can manufacture iPhones now.
But sure, China's the obvious winner this time. Good luck.
Re: Kimi K3, Qwen 3.8, and Anthropic's (Potential) Unravelling
#315The open weight, open architecture releases of the past several days has me more convinced that ultimately, the winner will be whoever burns their models to ASICs fastest. The LLMs themselves are capable of doing some aspects of chip design as evinced by the K3 press release. Furthermore, the frontier models are "good enough" for a wide swathe of tasks and will soon hit that threshold for a good amount of software en…
> whoever burns their models to ASICs fastest. There is already custom hardware see cerebras. GPUs have a lot of slack there is at least one lab that had a (small 8b) model generate almost 3000 tokens per second on a MI300X for a talk, instead of the typical software stack that did maybe 100ish tokens per second. High bandwidth flash storage is in the works, i.e hard drives with TBs of storage and over 1 TB per secon…
There is no "may" here. You will see this.
It's always difficult to see it from the present, but we're not at some end stage in hardware development; we're still on the same curve our predecessors also couldn't see: they couldn't imagine that there would be high performance computers carried in our pockets, with staggering amounts of storage and compute, putting to shame the machines they filled rooms with.
Re: Kimi K3, Qwen 3.8, and Anthropic's (Potential) Unravelling
#316I keep thinking about the Figma thing. If you're unaware, here's the google summary: ---- The Board Departure: Mike Krieger, Anthropic’s CPO and a co-founder of Instagram, sat on Figma’s board of directors. He resigned on April 14, just days before news of Claude Design broke. This sparked speculation over conflict of interest and the use of proprietary product strategy information. Betrayal of Partnership: The launc…
Your last two sentences remind me of the olden days of people building applications on top of FB and Twitter APIs (and later, Reddit), only to to have the rug pulled out from under them in one way or another
story as old as internet itself!
Re: Kimi K3, Qwen 3.8, and Anthropic's (Potential) Unravelling
#317Earlier quoted context omitted.
I am really confused about the point you're trying to make. 150tok/s is slow, but so is 9000tok/s? Or they're both fast? Or 150tok/s should be enough?
The "640k should be enough for anyone" quote (even if Gates didn't exactly say it) is making the point is that _right now_ we have no idea about what our future needs and capabilities will be, we can't imagine what "should be enough" will be 640k was enough ... in 1981 ... almost fifty years later is 50,000 lower than a standard off the shelf PC now
The comment seems like the nerd equivalent of 6 7
Re: Kimi K3, Qwen 3.8, and Anthropic's (Potential) Unravelling
#318Upper bound of AI progress - recursive self improvement. In this case AI will be responsible for building better models, making people who own datacenters the winners. Anthropic/OAI is cooked. Lower bound of AI progress - plateau. Progess is slowing, focus is on serving a meaningful peak capability at the lowest possible price. There's been news today that Google is building a Gemini chip with weights baked into sili…
>. In this case AI will be responsible for building better models, making people who own datacenters the winners. We need SETI@home for Open Weight models yesterday...
Re: Kimi K3, Qwen 3.8, and Anthropic's (Potential) Unravelling
#319Earlier quoted context omitted.
Sunk cost fallacy is not relevant to future investment.
what an absolute plonker do you know how ROIC is calculated? Good luck hiding your huge sunk cost of bilions and billions in there bro. I wish people who had zero understanding of actual finance would never comment about it. Moreover if they declare their existing assets are bunk, the valuation is marked down, especially after now pricing in failure risk - this is catastrophic for VC's. Again, bro, just be quiet.
Re: Kimi K3, Qwen 3.8, and Anthropic's (Potential) Unravelling
#320I think the risk is overstated. For one, on the margin people are willing to pay a lot for slightly better models. I know personally the value the LLM adds to my workflow is considerably more than the $200/m I pay the frontier labs. I have no interest in optimizing that to get it slightly lower. There are a very vocal minority that optimizes this or companies whose LLM expense is marginal, but I think that's the mino…
Why should I pay for a frontier model that I can’t use?