
Alibaba's Qwen team has released open-weights versions of Qwen3.8, including a 27-billion-parameter model that outperforms Qwen3.7-Plus on coding and office tasks despite being smaller.
Both models are licensed under Apache 2.0 and available now on Hugging Face and ModelScope, with a cloud-hosted version coming soon to Qwen Cloud.
What happened
Alibaba's Qwen team released open weights for Qwen3.8-27B, a multimodal model with 27 billion parameters under the Apache 2.0 license. The model outperforms the larger Qwen3.7-Plus in coding and office tasks, handles up to 262,000 tokens of context natively (scaling to one million via YaRN), and processes text, images, and video. A larger Qwen3.8-2.4T-A95B model was also released, and both are available on Hugging Face and ModelScope.
Why it matters
The open-weights release under a permissive license lets developers and businesses freely use and modify the model for their own applications without licensing restrictions. The model's multimodal capabilities and improved agent autonomy expand what open-source developers can build; companies interested in coding or office automation tasks gain an alternative to proprietary systems.
What to watch
Qwen Cloud, Alibaba's AI service, will soon offer a hosted version with one million tokens of context, giving businesses a managed option for handling very long documents and conversation histories.
Alibaba's Qwen team released open-weights models for Qwen3.8, marking a significant open-source AI release. The flagship model, Qwen3.8-27B, is a multimodal dense model with 27 billion parameters. According to Qwen, it outperforms the larger Qwen3.7-Plus in coding and office tasks—a notable achievement given its smaller size. The team also highlighted improved agent capabilities, with the model planning more independently and completing tasks more reliably.
The Qwen3.8-27B natively handles up to 262,000 tokens of context and can scale to one million tokens using the YaRN method, allowing it to process very long documents and multi-turn conversations. Beyond text, the model is multimodal and processes images, videos, diagrams, documents, and multi-hour video content. A flexible thinking mode is enabled by default but can be toggled on a per-query basis, giving users control over the model's reasoning depth.
Qwen also released weights for Qwen3.8-2.4T-A95B, a much larger model built for Max-level operation. Both models are distributed under the Apache 2.0 license, a permissive open-source license that permits free use and modification. The models are available immediately on Hugging Face and ModelScope. To serve users who prefer managed deployment, Qwen announced that a hosted version with one million tokens of context will soon be available through Qwen Cloud, Alibaba's AI service, offering an alternative for teams that want the model's capabilities without self-hosting infrastructure.
Alibaba's Qwen team is positioning the open-weights release as a competitive move in the growing market for accessible AI models. By releasing a 27-billion-parameter model that outperforms a larger proprietary version (Qwen3.7-Plus) on specific tasks like coding and office automation, Qwen demonstrates that parameter count alone does not determine performance—specialized optimization can achieve better results with fewer resources. The Apache 2.0 license is notably permissive, removing licensing friction for developers and enterprises building on top of the model.
The inclusion of multimodal capabilities (text, images, and video) and the default flexible thinking mode—which can be toggled per query—suggests Qwen is targeting use cases that require reasoning and visual understanding. The planned cloud-hosted version through Qwen Cloud, with one million tokens of context, addresses a key bottleneck for production systems: handling extremely long documents and conversation histories without expensive custom deployment. This two-tier strategy (open-source for full control, cloud-hosted for convenience) is designed to capture both developer and enterprise segments.
AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.
Free · takes 30 seconds · unsubscribe anytime
Ask AI anything about this article. Q&As are published on this page for other readers too.
MCPMark V2 benchmarks showed that production LLMs achieve 54% higher token usage when reasoning through backen…

Researchers tested whether frontier AI models can conduct independent research by having Claude Opus 4.8 with…

OpenAI launched Computer History, a macOS feature that records user clicks, keystrokes, keyboard shortcuts, an…

Google now allows users to toggle off visible watermarks (the "sparkle" that appears in the bottom-right corne…

OpenAI is previewing Ultrafast mode, powered by Cerebras infrastructure from a ten-billion-dollar partnership…

Snowflake's Observe announced general availability of a redesigned MCP server and new CLI tool that give AI ag…

The AI news that matters, in one minute each morning.
Sign up free