Qwen3.8 Flash Next
MultimodalQwen3.8 Flash Next is an open-weight multimodal MoE from Alibaba's Qwen team, released August 2026 as an early preview of the Qwen4 architecture. A 125B-parameter backbone with 51B N-gram embeddings and a 4B multi-token prediction module activates only 6B parameters per token. It introduces a GDN + Qwen Sparse Attention hybrid, Gated Residual and N-gram Embedding, and reportedly costs about one-ninth of Qwen3.7-Plus to train. Native 262K context, extensible to 1M with YaRN. Licensed under the Qwen community license (not Apache-2.0).
Key Specifications
Timeline
Technical Specifications
Benchmark Results
Model performance metrics across various tests and benchmarks
Reasoning
Other Tests
License & Metadata
Compare Qwen3.8 Flash Next
All comparisonsArticles about Qwen3.8 Flash Next
Similar Models
All ModelsQwen3.5-397B-A17B
Alibaba
Qwen3.8 Max
Alibaba
Qwen3 235B A22B
Alibaba
Step-3.5-Flash
StepFun
Kimi K3
Moonshot AI
Kimi K2.5
Moonshot AI
Qwen3 VL 4B Instruct
Alibaba
Qwen3 VL 4B Thinking
Alibaba
Recommendations are based on similarity of characteristics: developer organization, multimodality, parameter size, and benchmark performance. Choose a model to compare or go to the full catalog to browse all available AI models.
