71 comments

[ 81.6 ms ] story [ 1556 ms ] thread
Are these latest Qwen models still open weights or has Qwen moved away from that?
Lmao I love their video with the idea that people will be able to do their hobbies while ai does their job.

Surely Alibaba is leading by example here by reducing work hours per week while keeping pay the same right? Right?

> Today, we are officially releasing Qwen 3.8-Max, the most capable model in the Qwen family to date. This also marks the first time we will open-source the weights of a Qwen-Max-class model — the open weights will be released next week.

I don't understand. That's dated today, but:

https://twitter.com/alibaba_qwen/status/2078759124914098291

> Qwen3.8 is launching and going open-weight soon! [...] You don't have to wait to test it. Just now, the Qwen3.8-Max-Preview made its debut on Alibaba’s Token Plan, Qoder, and QoderWork.

That was on July 19th. I used it to draw this pelican: https://simonwillison.net/2026/Jul/20/afraid-of-chinese-mode...

So what are they releasing today?

They've also announced Qwen3.8-27B being released open-weight next week. Qwen3.6-27B is widely regarded as one of the best local models, especially since nothing else comes close to it, that isn't benchmaxxed, without being significantly larger. If 3.8 truly improves upon it that would be awesome.
It is really a difficult choice in front of me: DeepSeek v4 flash 0731@q2 vs Qwen 3.{6,8} 27B@fp8 on a 96GB VRAM server.
It seems this is the only mention of cost?

> Qwen3.8-Max comes with the official support for reasoning_effort, which can be used to adjust reasoning depth and control cost:

> xhigh (default): for complex tasks demanding thorough analysis

> medium: balancing accuracy and speed

> low: efficient reasoning optimizing for speed and cost

I hope this is significantly cheaper. I've been loving Deepseek for it's nearly free usage costs, hard to justify switching from cents per day.

> This also marks the first time we will open-source the weights of a Qwen-Max-class model — the open weights will be released next week.

Nice!

“self-evolves through feedback loops”

Does this mean they distilled Claude? Sounds like what Claude Code will often do.

We will eventually need a self evolution benchmark to see where these large models can create recursive solutions that improve
Once OpenAI and Anthropic are public, every such announcement will become a reliable sell signal
the benchmark I trust most is whether the model can explain its own pricing page without getting confused
> In this case, Qwen3.8-Max was asked to create the oh-my-cli project from scratch and, over a 10+ day long-horizon autonomous coding run, build a self-evolving harness.

They don't explain how successful that went but it's a bit hilarious seen that an Anthropic dev explained that it's been 15 days Claude was hard at work --with nothing to show yet-- trying to rewrite itself in another language.

"You rewrite Claude Code, we rewrite oh-my-pi."

"You're nowhere after 15 days, we do it in 10."

Sure, it's apples to oranges and all that. But part of me thinks they know fully well what they did there.

Can a model be stripped off anything not relevant to coding and get a lot lighter? Or is that impossible?

Just like we have professors with specialisation wondering if AI models can also be so.

Does the page actually load for anyone? I get stupid spa skeleton spinners.
I think the window for a ban of open weight models is closing fast so let's hope US administration is going to miss it and we get Fable-level models (at least in some aspects) with open weights without infringing any newly introduced law as a long-term local baseline.
Does anyone know how token- and reasoning efficient it is? The charts don't show how many tokens were used in any benchmark.
I'm trying and failing to find value running a potential Qwen 3.8 27b dense model on a 16 core, 128 GB of ram, 2080ti box. Yes, the GPU yells for help, but the problem is that no math works to upgrade this machine even when pouring $200 in rent every month into the large model providers...

How are you all justifying economical use of these local models right now? What's the cost efficient way to do this and do better (even with models evolving over time and losing now vs later) than the big labs?

is it the right time to perhaps switch to QwenCode ?

i might end up cancelling claude, anybody else thinking of the same ?

Tokenpocalypse canceled
Has anyone tried Qwen with the Fusion 360 MCP server? I feel like drawing with python is close enough but I'm curious
Is "cowork" a general industry term now? Here I was just getting used to "coding" replacing "programming".
It was a matter of time for China to catch up with the US. In terms of infrastructure, manufacturing, and engineering workforce, China has the upperhand and I foresee them becoming the SOTA leaders. Maybe if the US wasn't so busy gatekeeping and keeping things proprietary, they would've had more trust from the open source community.