Just imagine Anthropic making Opus open-weights now for the sake of trolling everyone. Wouldn't surprise me at this point xD
Has Anthropic released anything at all, ever?
Qwen 3.8
501–510 of 793 posts
Re: Qwen 3.8
#502Earlier quoted context omitted.
Anyone else think the AI environmental backlash is astroturfed? I keep looking at the numbers. The power use numbers are not that problematic. Ordering a burrito on DoorDash uses more power than a few days of heavy AI use. The water argument applies to some locations, and is mostly a local governance problem... if the data centers are using too much water, it means they are not being charged enough for that water. Ch…
> Yet the visceral pile-on here is so extreme, it feels fake. Driven by people in the few roles that are soundly replaced by AI-- e.g. low tier media slop producers, who hate AI because it threatens their socially negative worthless jobs. The arguments are so paper thin because the environmental impact isn't their concern, it's just a target that sounds convincing to people who don't know better.
Historically art of any kind is a U-shaped market: there is low-end work and high-end work. Nothing in between.
So I do understand some of the AI hate among that population. It's chopping the bottom tier work off. Either you're a top-tier massively successful artist or there is $0 to be made anywhere doing anything.
Long term I think it will do that to all white collar work. There will be no entry level jobs. Period. None. Zero. You're either very experienced or there is no work.
This is a huge problem, and one we will have to address.
Re: Qwen 3.8
#503"Qwen3.8 is launching and going open-weight soon! With a massive 2.4T parameters..."
Re: Qwen 3.8
#504Earlier quoted context omitted.
I can't really blame them that the biggest labs focused on trainig and realeasing huge models. The niche for small models should be filled with medium sized labs doing distillations of the huge ones into consumer grade hardware runnable models and LORAs for the huge ones.
I think AI will evolve the same way computers did. We're somewhere in the 80s-90s timeline of the evolution. My prediction is that on-device models will have excellent tool-calling, reasoning, and general skills, but the domain-specific knowledge will be retrieved on-demand from vendors like Google. Rather than downloading models, each device will have a hardware component with weights baked into silicon for maximum…
Re: Qwen 3.8
#505It feels like an inflection point of lost US leadership in technology? A year plus ago you would say while China led in green energy and manufacturing, at least the US was ahead in software - as demonstrated by the state of US AI models. We could point at a lot of factors on the US side. From political paralysis / head-in-the-sand attitudes towards emerging tech like green energy. To something of disdain for workers…
It's just lack of antitrust enforcement.
China pours money into tons of different businesses in the same industry and lets them fight it out. The only US business model left now is to shut down (or collude with) competitors and raise prices while cutting costs. All they have to do is cut Congress (and regulators, and individual judges) in. We've financialized everything for the sake of scammers, rather that finance being used for the sake of getting cash to the most productive organizations. We've optimized for corruption.
If we hadn't let the stupidest people in the world buy up everything, and made doing nothing with it the most profitable option, China would have never have blown past us.
The US Supreme Court has explicitly legalized "tipping" politicians. That's the biggest sign of degeneracy that a government could possibly achieve.
Re: Qwen 3.8
#506Earlier quoted context omitted.
But they didn't. The deal to buy 40% of the world's memory never happened after the price increases the news generated.
How dare they not buy things
Re: Qwen 3.8
#507Earlier quoted context omitted.
> How does this explain open weights? They could easily take the same closed route like their American friends Because they are playing the Americans at their own game. What is the first thing an American company would do ? Spread the old American classic FUD ... "you can't used this closed tool because its run by the communists", right ? So you release it as open weights which is a win-win. Global adoption of the mo…
I think most of us that will claim to understand China are going to end up being wrong, unless any of us live there or grow up there. There’s a saying about China I have heard from ex-pats: the more you know about China, the less you know about China. The point of me bringing that up is to say that what follows is really just my best guess: If I were to judge from China’s approach to hardware, I think that the compan…
That is a very thin moat, though. There's nothing you can do with, for example, Claude Code + Opus 4.8 that you can't do with your own custom harness running API-level Opus 4.8, which means that if you can afford the hardware (the moat for running any SOTA model) you don't need to pay Anthropic anymore.
I'm not saying they shouldn't, but I understand why they don't.
Re: Qwen 3.8
#508Things are heating up in China. Looking forward to see what Antrophic and OpenAI does next.
Re: Qwen 3.8
#509I assume that this announcement has been prompted by that of Moonshot AI, which has just announced a 2.8T parameter open-weights LLM, Kimi K3, to be published on Huggingface by 27 July. Now the response of Alibaba is that they will also publish soon a big open weights LLM, the 2.4T parameter Qwen 3.8. I wonder if Alibaba has always planned to make this big LLM open weights, or they have chosen to do this now, to bett…
It's important to note there was recently a large AI conference in Shanghai, and Xi Jinping mentioned a commitment to open source AI releases. It is no surprise that Alibaba would want to align.
Re: Qwen 3.8
#510Earlier quoted context omitted.
DeepSeek V4 pricing is insane, 10x-30x cheaper to use than most other models, and it usually is good enough for most tasks.
> it usually is good enough for most tasks The model is fantastic. And costs almost nothing. The only problem I see is that they will train on your data. There are zero-data-retention providers of DeepSeek models, of which I have used openrouter (with zdr guardrails), and fireworks. But these are 3x to 5x more expensive than directly using DeepSeek, possibly due to poor caching. Thats the price to pay for zdr.