Qwen3.8

Alibaba's 2026 open-weight flagship family, from a 2.4T MoE down to an Apache-licensed 27B model.

Open SourceSelf HostedOffline CapableGPU Required
0.0 (0)

About

Alibaba's August 2026 generation spans the widest range of any open family. At the top, Qwen3.8-2.4T-A95B, the open-weight release of the hosted Qwen3.8-Max, is second in size only to Kimi K3 among open models: 2.4 trillion total parameters with 95B active, built from a hybrid stack that interleaves Gated DeltaNet linear-attention blocks with gated full attention, routing through 512 experts per MoE layer. It runs exclusively in thinking mode, handles text only, reads 262K tokens natively with extension toward 1M, and downloads from Hugging Face under the custom Qwen3.8-Max license. At the other end, Qwen3.8-27B arrived in mid-August 2026 under Apache 2.0 as a dense vision-language model that reads images and video and posts strong coding scores, with 4-bit quantized builds around 17GB that fit a single consumer GPU. Recommended serving for the big checkpoints is SGLang, vLLM, or TokenSpeed on multi-GPU clusters, with Alibaba Cloud offering hosted endpoints; the 27B's permissive terms make it the practical choice for commercial local deployment.

Reviews (0)

Leave a Review

No reviews yet. Be the first to review!

Details

Price
Freemium
Platform
Hybrid
Difficulty
Advanced (4/5)
License
Qwen3.8-Max License
Added
Aug 24, 2026

Related Tools

Featured

Lightweight open-weight LLM by Google available in 1B to 27B sizes.

Open SourceSelf HostedOfflineGPU 4GB+
Easy

Open-source code LLM family by IBM for enterprise code generation.

Open SourceSelf HostedOfflineGPU 8GB+
Intermediate

Open-weight LLM by Meta in 8B and 70B sizes with strong general capabilities.

Open SourceSelf HostedOfflineGPU 8GB+
Intermediate
Featured

High-performance open-weight MoE LLM with 671B total parameters.

Open SourceSelf HostedOfflineGPU 24GB+
Advanced

Hybrid SSM-Transformer model by AI21 Labs combining Mamba with attention layers.

Open SourceSelf HostedOfflineGPU 12GB+
Intermediate

Open-weight code LLM trained on 2 trillion tokens of code and natural language.

Open SourceSelf HostedOfflineGPU 8GB+
Intermediate
Browse all Large Language Models (LLMs) tools