Alibaba open-weights Qwen 3.8 with a 27B model that tops its bigger predecessor

Alibaba's Qwen team published the open weights for Qwen3.8 on Thursday, releasing the models under the Apache 2.0 license. The move puts a capable new open model directly into the hands of researchers and companies that want to fine-tune and self-host it. The centerpiece is Qwen3.8-27B, a 27 billion parameter multimodal dense model. Qwen says it outperforms the larger Qwen3.7-Plus on coding and office tasks, a notable result for a model less than a tenth the size of some frontier flagships. The

2 min
Alibaba open-weights Qwen 3.8 with a 27B model that tops its bigger predecessor

Alibaba's Qwen team published the open weights for Qwen3.8 on Thursday, releasing the models under the Apache 2.0 license. The move puts a capable new open model directly into the hands of researchers and companies that want to fine-tune and self-host it.

The centerpiece is Qwen3.8-27B, a 27 billion parameter multimodal dense model. Qwen says it outperforms the larger Qwen3.7-Plus on coding and office tasks, a notable result for a model less than a tenth the size of some frontier flagships. The team also highlights stronger agent behavior, with the model planning more independently and completing multi-step tasks more reliably.

qwen3-8-open-weights-released

Qwen3.8-27B handles up to 262,000 tokens of context natively and can stretch to one million using the YaRN method. It reads images and video, including diagrams, documents, and multi-hour recordings, and ships with a flexible thinking mode that is on by default but can be toggled per query.

Alongside the 27B release, Qwen published weights for Qwen3.8-2.4T-A95B, a much larger model built to the Max tier. Both are available on Hugging Face and ModelScope, and a hosted version with one million tokens of context is expected soon on Qwen Cloud, Alibaba's AI service.

The open release follows Alibaba's Aug 3 launch of the Qwen 3.8-Max flagship, whose early benchmarks drew scrutiny. By open-weighting the 3.8 family under Apache 2.0, Alibaba gives the broader AI community a transparent, modifiable baseline rather than a closed API.

Sources

The Decoder, "Alibaba's Qwen team releases Qwen 3.8 models with open weights under the Apache 2.0 license" (Aug 14, 2026): https://the-decoder.com/alibabas-qwen-team-releases-qwen-3-8-models-with-open-weights-under-the-apache-2-0-license/

Qwen via X: https://x.com/Alibaba_Qwen/status/2088280182356611304

Hugging Face, Qwen3.8 collection: https://huggingface.co/collections/Qwen/qwen38

Techmeme aggregation: https://www.techmeme.com/260814/p16#a260814p16

Written by

More to read

  • Anthropic Demonstrates Autonomous De Novo Protein Design and Chemical Analysis with Claude

    Anthropic Demonstrates Autonomous De Novo Protein Design and Chemical Analysis with Claude Anthropic has published experimental results demonstrating Claude's ability to autonomously design de novo protein binders with physical wet-lab validation and automate complex analytical chemistry workflows. The findings show frontier LLMs acting as autonomous agents across computational biology and molecular characterization pipelines. In the primary experiment, Anthropic evaluated Claude Mythos Previe

    1 min
  • Cerebras Unveils CS-4 Rack-Scale System Powered by Three WSE-3 Turbo Chips and Nexus Architecture

    Cerebras Unveils CS-4 Rack-Scale System Powered by Three WSE-3 Turbo Chips and Nexus Architecture Cerebras Systems has announced the CS-4, a rack-scale AI accelerator system designed around three of its next-generation Wafer Scale Engine 3 Turbo (WSE-3 Turbo) chips and a modular hardware architecture dubbed Nexus. Cerebras confirmed that initial customer shipments for the CS-4 are scheduled to begin in the current quarter. The new system marks a structural shift from Cerebras's single-wafer CS

    1 min
  • AI FinOps: Cutting LLM Inference Costs by 30-60% Through Model Tiering, Caching, and GPU Optimization

    AI FinOps: Cutting LLM Inference Costs by 30-60% Through Model Tiering, Caching, and GPU Optimization Inference costs have become the second-largest line item in enterprise AI budgets, trailing only talent spend according to RapidData's State of Enterprise AI 2026. This shift represents a fundamental inversion from the 2021-2023 era when training dominated AI expenditure. The compounding nature of serving costs—accumulating every hour as long as users hit the API—means that even modest producti

    1 min