Qwen3.8: A 2.4T-Parameter Open Model That Rivals Opus 4.8 on Coding Benchmarks
Qwen3.8-2.4T
Qwen releases Qwen3.8-2.4T-A95B, a massive mixture-of-experts model with 2.4 trillion total parameters and 95 billion activated per token. It introduces a hybrid architecture combining gated DeltaNet and attention layers, supports a native 262K context (extendable to 1M), and features adjustable reasoning effort. Benchmark results show it competing with top proprietary models like Claude Opus 4.8 and GPT-5.6 Sol on coding and agentic tasks, with FP8 quantized weights available for efficient deployment via vLLM, SGLang, and TokenSpeed.
Following the widespread community adoption of the Qwen3.5 and Qwen3.6 series, we are pleased to introduce Qwen3.8, the most capable generation in the Qwen open-model family to date.