Alibaba logo

Qwen3 235B A22B

Alibaba

Qwen3 235B A22B is a large language model from Alibaba with a Mixture-of-Experts (MoE) architecture, containing 235 billion total parameters and 22 billion active parameters. Achieves competitive results on coding, math, general capabilities, and other benchmarks compared to other top models.

Key Specifications

Parameters
235.0B
Context
128.0K
Release Date
April 28, 2025
Average Score
76.2%

Timeline

Key dates in the model's history
Announcement / Last Update
April 28, 2025
Today
October 5, 2026

Technical Specifications

Parameters
235.0B
Training Tokens
36.0T tokens
Knowledge Cutoff
-
Family
-
Capabilities
MultimodalZeroEval

Pricing & Availability

Input (per 1M tokens)
$0.20
Output (per 1M tokens)
$0.60
Max Input Tokens
128.0K
Max Output Tokens
128.0K
Supported Features
Function CallingStructured OutputCode ExecutionWeb SearchBatch InferenceFine-tuning

Benchmark Results

Model performance metrics across various tests and benchmarks

General Knowledge

Tests on general knowledge and understanding
MMLU
Accuracy • Self-reported
88.0%

Programming

Programming skills tests
MBPP
Accuracy • Self-reported
81.4%

Mathematics

Mathematical problems and computations
GSM8k
Accuracy • Self-reported
94.0%
MATH
Accuracy • Self-reported
71.8%
MGSM
Accuracy • Self-reported
83.5%

Reasoning

Logical reasoning and analysis
GPQA
Accuracy • Self-reported
47.5%

Other Tests

Specialized benchmarks
Arena Hard
Accuracy • Self-reported
96.0%
BBH
Accuracy • Self-reported
89.0%
MMMLU
Accuracy • Self-reported
87.0%
Aider
Pass@2 • Self-reported
61.8%
AIME 2024
Pass@64 • Self-reported
85.7%
AIME 2025
Pass@64 • Self-reported
81.5%
BFCL
v3 • Self-reported
70.8%
CRUX-O
Score • Self-reported
79.0%
EvalPlus
Score • Self-reported
77.6%
Include
Score • Self-reported
73.5%
LiveBench
Accuracy • Self-reported
77.1%
LiveCodeBench
v5 • Self-reported
70.7%
MMLU-Pro
Accuracy • Self-reported
68.2%
MMLU-Redux
Accuracy • Self-reported
87.4%
MultiLF
Accuracy • Self-reported
71.9%
MultiPL-E
Score • Self-reported
65.9%
SuperGPQA
Accuracy • Self-reported
44.1%

License & Metadata

License
apache-2.0
Announcement Date
April 28, 2025
Last Updated
April 28, 2025

Articles about Qwen3 235B A22B

Similar Models

All Models

Recommendations are based on similarity of characteristics: developer organization, multimodality, parameter size, and benchmark performance. Choose a model to compare or go to the full catalog to browse all available AI models.