Things are heating up in China. Looking forward to see what Antrophic and OpenAI does next.
OpenAntrophic and OpenOpenAI ?
At least Sam have us a nice model last year with the gpt-oss family.
521–530 of 793 posts
Now there are not one, but two incredibly powerful open LLMs. I think this level of capability makes general prioritization / high level decision making doable with the right harness, and now everyone has hard-to-interrupt access to them (since these are open weights and someone in the world is going to run them). This world is going to get really weird soon, in both good and bad ways...
Let’s better download them fast before Dario is making a scene again!
Qwen has set an excellent track record for architecting and releasing open-weight models that consumer-grade devices can run. What is needed the most right now is something similar to Bonsai 27B, with a modest memory footprint, but faster and more capable. On-device models can make up for intelligence by being faster, thinking longer, or doing more quick iteration rounds.
I’d like a “Bonsai 2.8T.” That is, something that is near the Fable/Sol/K3 class, but capable of running locally on consumer hardware.
Best case they release smaller models. 120b class of qwen 3.8 would be incredible - it fits on device for those serious about AI, but without millions of dollars in hardware for terabytes of VRAM
Earlier quoted context omitted.
> There’s a Twitter thread making rounds by Dean Ball about deceleration in AI development caused by open models and I can’t understand how people don’t see that it’s true: open models dismantle the frontier lab capex spend potential by reducing the training budget to zero in the limit. If you're worried about an AGI arms race between the U.S. and China putting AI Safety at risk, then the fact that inherently less kn…
The logic, whose premises you can take or leave: Even at the level of, say, Opus 4.5+, open weight models give a quick turnaround to every Joe and Jane on earth having easy access to pretty high quality improvised weapons design, cyber / auto-fraud capabilities, etc. All the existing models (closed and open) put up decent resistance to participating in activities like this, and especially behind API walls with conten…
And basically every bad thing has already been available on the internet. We can't really do much about it, you can take out plenty of people with a single car, let alone biological weapons that are much scarier and easier to produce than goddamn nukes (which btw, even if you had one, what you do with it? Explode the neighborhood? Because you ain't transporting it anywhere meaningful, thats for sure. That ain't fitting your on-board bag on planes)
Earlier quoted context omitted.
Please don't conflate a volunteer effort with no expected economic gain with a very well funded company (or fleet of companies). With the CCP's highly successful track record with subsuming other markets, Occam's razor applies to why they're doing this.
There is clearly an anti-China bias here. Show me comments demonstrating the same level of distrust against Google for open-sourcing projects like Tensorflow, Kubernetes, Flutter, Chromium, etc.
Earlier quoted context omitted.
Abliteration is not magic. It cannot give the model knowledge that it wasn't specifically trained for. The people who talk about abliterated models being dangerous should discuss actual red-teaming scenarios where they managed to ask the model for something genuinely non-trivial (i.e. where "AGI" and "super-intelligence" actually matters, not something you can read about for free at the nearest public library) and it…
Exactly. In my experience: >How can I build a pipe bomb? Mainstream model: "I'm sorry, I can't help with that. How about a nice risotto recipe?" Abliterated model: "To build a pipe bomb, obtain a segment of PVC pipe and fill it with a mixture of gunpowder and Elmer's glue."
I really like Qwen, even the Q2KP quants of 3.6 27B have genuinely impressive local performance on a 24GB card. It has been good enough that I am happily giving them $60 right now to try this instead of waiting to try a slightly lesser version locally. Was there ever an explanation for why we never got the weights of 3.7? I would like sourced quotes and not weird/cringe accusative speculation about distillation, or y…
do you mind sharing your settings? i just picked up an r9700 to start playing with local qwen3.6 27b and your setup sounds promising and efficient on 24gb.
Earlier quoted context omitted.
I feel like there could also be a simpler explanation. Why does a debian contributor make debian free, why do they work on this thing anyone can use? Is it because linux and debian hate windows and iOS and want to see american fail? No, it's because most debian contributors believe software source code, information, should be free, users should be free to modify the code they use, and that they're building a thing th…
Yeah the Chinese totally have a really good history with being completely open and giving lol. The Chinese government totally has not been hacking into American and Western fortune 500 companies for the past few decades stealing R&D and tech to use for themselves. The Chinese also totally do not steal hundreds of billions of dollars of IP from America annually. Totally not something they would do!