Qwen3-Next-80B-A3B-Thinking
Qwen3-Next-80B-A3B-Thinking is the reasoning variant of the Qwen3-Next series from Alibaba's Qwen team, released 2025-09-10 under Apache 2.0 with open weights. It shares the instruct model's hybrid attention stack (Gated DeltaNet plus Gated Attention for ultra-long context) and a high-sparsity Mixture-of-Experts layout with 80B total and 3B active parameters, and was post-trained with GSPO to stabilise reinforcement learning over that hybrid architecture. It uses roughly 15T training tokens and is served by Novita with a 65K-token context window.
Key Specifications
Timeline
Technical Specifications
Pricing & Availability
Benchmark Results
Model performance metrics across various tests and benchmarks
Reasoning
Other Tests
License & Metadata
Similar Models
All ModelsQwQ-32B-Preview
Alibaba
Qwen2.5 72B Instruct
Alibaba
QwQ-32B
Alibaba
Qwen2.5 32B Instruct
Alibaba
Qwen2 72B Instruct
Alibaba
Qwen3 30B A3B
Alibaba
Qwen2.5 14B Instruct
Alibaba
Qwen3 32B
Alibaba
Recommendations are based on similarity of characteristics: developer organization, multimodality, parameter size, and benchmark performance. Choose a model to compare or go to the full catalog to browse all available AI models.