Earlier quoted context omitted.
I have a 128G mac studio and even 397B was a happy surprise to me due to its high quantization resilience. I've created a 2.54BPW quant that fit on my hardware with 128k context, 20 tps tg and 200tps pp, while maintaining high scores on many benchmarks: https://huggingface.co/tarruda/Qwen3.5-397B-A17B-GGUF/discus...
Apple store's current options for mac studio seem to max out at 96GB. I'm questioning ROI, esp. given it's not upgradeable. Curious about others' takes on new mac hardware.
Qwen3.7-Max: The Agent Frontier
51–60 of 317 posts
Re: Qwen3.7-Max: The Agent Frontier
#52Earlier quoted context omitted.
Unless US hyperscalers do the same in reverse, I hope the status quo stays as it is. Either people are happy to share, and the sharing should happen both ways, or US hyperscalers can keep isolating themselves as they've done so far.
I do hope The U.S. hyperscalers do the same as well. In an ideal world U.S. residents would use Chinese AI models and Chinese residents would use U.S. AI models. Governments in both countries are collecting data for nefarious reasons. But the Chinese government has far less influence on a U.S. resident and vice versa. We are all better off if our data is collected by a government halfway across the world instead of o…
Re: Qwen3.7-Max: The Agent Frontier
#53Re: Qwen3.7-Max: The Agent Frontier
#54Earlier quoted context omitted.
Unless US hyperscalers do the same in reverse, I hope the status quo stays as it is. Either people are happy to share, and the sharing should happen both ways, or US hyperscalers can keep isolating themselves as they've done so far.
I do hope The U.S. hyperscalers do the same as well. In an ideal world U.S. residents would use Chinese AI models and Chinese residents would use U.S. AI models. Governments in both countries are collecting data for nefarious reasons. But the Chinese government has far less influence on a U.S. resident and vice versa. We are all better off if our data is collected by a government halfway across the world instead of o…
On the other hand, there's other models where the source is 100% open, the training data is known, and people have reproduced the same model from scratch, so while those trail behind, there's definitely an effort to make models more open and capable.
Re: Qwen3.7-Max: The Agent Frontier
#55Earlier quoted context omitted.
I do hope The U.S. hyperscalers do the same as well. In an ideal world U.S. residents would use Chinese AI models and Chinese residents would use U.S. AI models. Governments in both countries are collecting data for nefarious reasons. But the Chinese government has far less influence on a U.S. resident and vice versa. We are all better off if our data is collected by a government halfway across the world instead of o…
China is much more interested in waging a campaign against companies that represent the material of the future growth in productivity, exports, and prosperity of the US and her people, than learning about you as an individual. Unless of course you are a Chinese dissident living in the US.
Re: Qwen3.7-Max: The Agent Frontier
#56As they start to release more proprietary models, I so wish that they partnered with one of the major US hyperscalers to allow using these models through something US-domiciled. Totally understand why it may not be reasonable or in their best interest (and that the US is _absolutely_ not doing the same reflexively). But it would be lovely to be able to try these out on production workloads in earnest.
Re: Qwen3.7-Max: The Agent Frontier
#57As they start to release more proprietary models, I so wish that they partnered with one of the major US hyperscalers to allow using these models through something US-domiciled. Totally understand why it may not be reasonable or in their best interest (and that the US is _absolutely_ not doing the same reflexively). But it would be lovely to be able to try these out on production workloads in earnest.
Re: Qwen3.7-Max: The Agent Frontier
#58These are very good numbers. I still don’t get why they don’t compare against latest competitor versions in these posts, it’s not like we’re all not going to notice.
I think the argument is that trying to suggest that they’re close to N months from SOTA. Realistically I assume they hope readers don’t notice the fine details. The Qwen models are great for open weights but for every past release they haven’t performed as well as the benchmarks in my experience. They’re optimizing for benchmark numbers because they know it works.
The pool of people reading such articles while ignoring such details can't be big.
Re: Qwen3.7-Max: The Agent Frontier
#59Earlier quoted context omitted.
Expect to pay $4k-10k - Your RTX 6000 is closer to $10k now - Sparks are creeping into the $4-5k range - AMD Strix are ~3.5k - Apple depends on chipset and memory. Sweet spot would be 128gb M3 Ultra, probably $6-8k but admittedly haven't been tracking closely. New M5 might come in the fall. You can get a new 128gb M5 Max laptop for ~5-6k today. - a 4x3090 rig would take $5-6k Every platform has tradeoffs, but it's mo…
And for what? Spend 10-15k for the slopiest of slop code, non deterministic automations, and the ability to spawn an AI gf? This whole thing is really starting to remind me of the crypto hype phases of 2016-2018 when everyone thought their investment in GPUs was going to make them rich.
Re: Qwen3.7-Max: The Agent Frontier
#60Earlier quoted context omitted.
I think the argument is that trying to suggest that they’re close to N months from SOTA. Realistically I assume they hope readers don’t notice the fine details. The Qwen models are great for open weights but for every past release they haven’t performed as well as the benchmarks in my experience. They’re optimizing for benchmark numbers because they know it works.
> Realistically I assume they hope readers don’t notice the fine details. The pool of people reading such articles while ignoring such details can't be big.
On Hacker News I wonder if most people even opened the article at all most times.