Qwen3.8 Flash Next
MultimodalQwen3.8 Flash Next is an open-weight multimodal MoE from Alibaba's Qwen team, released August 2026 as an early preview of the Qwen4 architecture. A 125B-parameter backbone with 51B N-gram embeddings and a 4B multi-token prediction module activates only 6B parameters per token. It introduces a GDN + Qwen Sparse Attention hybrid, Gated Residual and N-gram Embedding, and reportedly costs about one-ninth of Qwen3.7-Plus to train. Native 262K context, extensible to 1M with YaRN. Licensed under the Qwen community license (not Apache-2.0).
Key Specifications
Timeline
Technical Specifications
Pricing & Availability
Benchmark Results
Model performance metrics across various tests and benchmarks
Reasoning
Other Tests
License & Metadata
Compare Qwen3.8 Flash Next
All comparisonsArticles about Qwen3.8 Flash Next
Similar Models
All ModelsQwen3.5-397B-A17B
Alibaba
Qwen3.8 Max
Alibaba
Qwen3.8 Flash
Alibaba
Qwen3.6 Plus
Alibaba
Qwen3.7-Plus
Alibaba
Qwen3.6-27B
Alibaba
Qwen3.6-35B-A3B
Alibaba
Qwen3.8-27B
Alibaba
Recommendations are based on similarity of characteristics: developer organization, multimodality, parameter size, and benchmark performance. Choose a model to compare or go to the full catalog to browse all available AI models.
