Alibaba logo

Qwen3-Next-80B-A3B-Thinking

Alibaba

Qwen3-Next-80B-A3B-Thinking is the reasoning variant of the Qwen3-Next series from Alibaba's Qwen team, released 2025-09-10 under Apache 2.0 with open weights. It shares the instruct model's hybrid attention stack (Gated DeltaNet plus Gated Attention for ultra-long context) and a high-sparsity Mixture-of-Experts layout with 80B total and 3B active parameters, and was post-trained with GSPO to stabilise reinforcement learning over that hybrid architecture. It uses roughly 15T training tokens and is served by Novita with a 65K-token context window.

Key Specifications

Parameters
80.0B
Context
65.5K
Release Date
September 10, 2025
Average Score
67.9%

Timeline

Key dates in the model's history
Announcement
September 10, 2025
Last Update
September 11, 2026
Today
October 5, 2026

Technical Specifications

Parameters
80.0B
Training Tokens
15.0T tokens
Knowledge Cutoff
-
Family
-
Capabilities
MultimodalZeroEval

Pricing & Availability

Input (per 1M tokens)
$0.15
Output (per 1M tokens)
$1.50
Max Input Tokens
65.5K
Max Output Tokens
65.5K
Supported Features
Function CallingStructured OutputCode ExecutionWeb SearchBatch InferenceFine-tuning

Benchmark Results

Model performance metrics across various tests and benchmarks

Reasoning

Logical reasoning and analysis
GPQA
• Self-reported
77.2%

Other Tests

Specialized benchmarks
AIME 2025
• Self-reported
87.8%
Arena-Hard v2
GPT-4.1 evaluated win rates • Self-reported
62.3%
BFCL-v3
• Self-reported
72.0%
CFEval
• Self-reported
20.7%
HMMT25
• Self-reported
73.9%
IFEval
• Self-reported
88.9%
Include
• Self-reported
78.9%
LiveBench 20241125
• Self-reported
76.6%
LiveCodeBench v6
25.02-25.05 • Self-reported
68.7%
MMLU-Pro
• Self-reported
82.7%
MMLU-ProX
• Self-reported
78.7%
MMLU-Redux
• Self-reported
92.5%
Multi-IF
• Self-reported
77.8%
OJBench
• Self-reported
29.7%
PolyMATH
• Self-reported
56.3%
SuperGPQA
• Self-reported
60.8%
Tau2 Airline
• Self-reported
60.5%
Tau2 Retail
• Self-reported
67.8%
Tau2 Telecom
• Self-reported
43.9%
TAU-bench Airline
• Self-reported
49.0%
TAU-bench Retail
• Self-reported
69.6%
WritingBench
• Self-reported
84.6%

License & Metadata

License
apache_2_0
Announcement Date
September 10, 2025
Last Updated
September 11, 2026

Similar Models

All Models

Recommendations are based on similarity of characteristics: developer organization, multimodality, parameter size, and benchmark performance. Choose a model to compare or go to the full catalog to browse all available AI models.