So, how large is that new model?
Qwen3-Max-Thinking
11–20 of 450 posts
Re: Qwen3-Max-Thinking
#12Re: Qwen3-Max-Thinking
#13I don't see a hugging face link, is Qwen no longer releasing their models?
afaiu not all of their models are open weight releases, this one so far is not open weight (?)
Re: Qwen3-Max-Thinking
#14Re: Qwen3-Max-Thinking
#15Earlier quoted context omitted.
afaiu not all of their models are open weight releases, this one so far is not open weight (?)
What would a good coding model to run on an M3 Pro (18GB) to get Codex like workflow and quality? Essentially, I am running out quick when using Codex-High on VSCode on the $20 ChatGPT plan and looking for cheaper / free alternatives (even if a little slower, but same quality). Any pointers?
If you had more like 200GB ram you might be able to run something like MiniMax M2.1 to get last-gen performance at something resembling usable speed - but it's still a far cry from codex on high.
Re: Qwen3-Max-Thinking
#16Earlier quoted context omitted.
afaiu not all of their models are open weight releases, this one so far is not open weight (?)
What would a good coding model to run on an M3 Pro (18GB) to get Codex like workflow and quality? Essentially, I am running out quick when using Codex-High on VSCode on the $20 ChatGPT plan and looking for cheaper / free alternatives (even if a little slower, but same quality). Any pointers?
Re: Qwen3-Max-Thinking
#17Earlier quoted context omitted.
afaiu not all of their models are open weight releases, this one so far is not open weight (?)
What would a good coding model to run on an M3 Pro (18GB) to get Codex like workflow and quality? Essentially, I am running out quick when using Codex-High on VSCode on the $20 ChatGPT plan and looking for cheaper / free alternatives (even if a little slower, but same quality). Any pointers?
The best could be GLN 4.7 Flash, and I doubt it's close to what you want.
Re: Qwen3-Max-Thinking
#18Mandatory pelican on bicycle: https://www.svgviewer.dev/s/U6nJNr1Z
Re: Qwen3-Max-Thinking
#19Earlier quoted context omitted.
afaiu not all of their models are open weight releases, this one so far is not open weight (?)
What would a good coding model to run on an M3 Pro (18GB) to get Codex like workflow and quality? Essentially, I am running out quick when using Codex-High on VSCode on the $20 ChatGPT plan and looking for cheaper / free alternatives (even if a little slower, but same quality). Any pointers?
If remote models are ok you could have a look at MiniMax M2.1 (minimax.io) or GLM from z.ai or Qwen3 Coder. You should be able to use all of these with your local openai app.
Re: Qwen3-Max-Thinking
#20Earlier quoted context omitted.
afaiu not all of their models are open weight releases, this one so far is not open weight (?)
What would a good coding model to run on an M3 Pro (18GB) to get Codex like workflow and quality? Essentially, I am running out quick when using Codex-High on VSCode on the $20 ChatGPT plan and looking for cheaper / free alternatives (even if a little slower, but same quality). Any pointers?
I gave one of the GPUs to my kid to play games on.