Qwen3.8 Flash
MultimodalQwen3.8 Flash is the production QwenCloud / OpenRouter API version of Qwen's 125B-parameter MoE Flash architecture, built on the open-weight Qwen3.8-Flash-Next checkpoint. It ships with a 1M-token default context, built-in tools, function calling, and structured output, with thinking enabled by default. Multimodal text, image, and video understanding at roughly one-ninth the training cost of Qwen3.7-Plus.
Key Specifications
Timeline
Technical Specifications
Pricing & Availability
Benchmark Results
Model performance metrics across various tests and benchmarks
Reasoning
Other Tests
License & Metadata
Compare Qwen3.8 Flash
All comparisonsSimilar Models
All ModelsQwen3.8 Flash Next
Alibaba
Qwen3.5-397B-A17B
Alibaba
Qwen3.8 Max
Alibaba
Qwen3.7-Plus
Alibaba
Qwen3.6-35B-A3B
Alibaba
Qwen3.8-27B
Alibaba
Qwen3 235B A22B
Alibaba
Inkling-Small
Thinking Machines Lab
Recommendations are based on similarity of characteristics: developer organization, multimodality, parameter size, and benchmark performance. Choose a model to compare or go to the full catalog to browse all available AI models.