Alibaba logo

Qwen3-Next-80B-A3B-Instruct

Alibaba

Qwen3-Next-80B-A3B-Instruct is the first model in the Qwen3-Next series with breakthrough architectural innovations. Uses hybrid attention (Gated DeltaNet + Gated Attention) for efficient ultra-long context modeling, MoE with high sparsity (512 experts, 10 active + 1 shared), and multi-token prediction. 80 billion parameters (3 billion active), trained on 15T tokens. Outperforms Qwen3-32B-Base at 10% of the training cost. Context support up to 256K (expandable to 1M with YaRN). Apache 2.0 license.

Key Specifications

Parameters
80.0B
Context
65.5K
Release Date
September 9, 2025
Average Score
67.0%

Timeline

Key dates in the model's history
Announcement
September 9, 2025
Last Update
February 12, 2026
Today
September 10, 2026

Technical Specifications

Parameters
80.0B
Training Tokens
15.0T tokens
Knowledge Cutoff
-
Family
-
Capabilities
MultimodalZeroEval

Pricing & Availability

Input (per 1M tokens)
$0.15
Output (per 1M tokens)
$1.50
Max Input Tokens
65.5K
Max Output Tokens
65.5K
Supported Features
Function CallingStructured OutputCode ExecutionWeb SearchBatch InferenceFine-tuning

Benchmark Results

Model performance metrics across various tests and benchmarks

Reasoning

Logical reasoning and analysis
GPQA
Self-reported by the model providerSelf-reported
72.9%

Other Tests

Specialized benchmarks
MMLU-Redux
Self-reported
91.0%
MultiPL-E
Self-reported
88.0%
IFEval
Self-reported
88.0%
WritingBench
Self-reported
87.0%
Creative Writing v3
Self-reported
85.0%
Arena-Hard v2
Evaluation GPT-4.1 by win rateSelf-reported
83.0%
Aider-Polyglot
Self-reported by the model providerSelf-reported
49.8%
AIME 2025
Self-reported by the model providerSelf-reported
69.5%
BFCL-v3
Self-reported by the model providerSelf-reported
70.3%
HMMT25
Self-reported by the model providerSelf-reported
54.1%
Include
Self-reported by the model providerSelf-reported
78.9%
LiveBench 20241125
Self-reported by the model providerSelf-reported
75.8%
LiveCodeBench v6
25.02-25.05Self-reported
56.6%
MMLU-Pro
Self-reported by the model providerSelf-reported
80.6%
MMLU-ProX
Self-reported by the model providerSelf-reported
76.7%
Multi-IF
Self-reported by the model providerSelf-reported
75.8%
PolyMATH
Self-reported by the model providerSelf-reported
45.9%
SuperGPQA
Self-reported by the model providerSelf-reported
58.8%
Tau2 Airline
Self-reported by the model providerSelf-reported
45.5%
Tau2 Retail
Self-reported by the model providerSelf-reported
57.3%
Tau2 Telecom
Self-reported by the model providerSelf-reported
13.2%
TAU-bench Airline
Self-reported by the model providerSelf-reported
44.0%
TAU-bench Retail
Self-reported by the model providerSelf-reported
60.9%

License & Metadata

License
apache-2.0
Announcement Date
September 9, 2025
Last Updated
February 12, 2026

Similar Models

All Models

Recommendations are based on similarity of characteristics: developer organization, multimodality, parameter size, and benchmark performance. Choose a model to compare or go to the full catalog to browse all available AI models.