Alibaba logo

Qwen3.8-27B

Multimodal
Alibaba

Qwen3.8-27B is a dense 27.78-billion-parameter multimodal foundation model for coding, professional work, research, and long-horizon agents.

Key Specifications

Parameters
27.8B
Context
-
Release Date
August 14, 2026
Average Score
66.3%

Timeline

Key dates in the model's history
Announcement
August 14, 2026
Last Update
August 29, 2026
Today
August 30, 2026

Technical Specifications

Parameters
27.8B
Training Tokens
-
Knowledge Cutoff
-
Family
-
Capabilities
MultimodalZeroEval

Benchmark Results

Model performance metrics across various tests and benchmarks

Reasoning

Logical reasoning and analysis
GPQA
Diamond splitSelf-reported
89.2%

Other Tests

Specialized benchmarks
Agents' Last Exam
ScoreSelf-reported
42.9%
AndroidWorld
Self-reported
81.9%
BabyVision
With code interpreterSelf-reported
85.6%
CharXiv-R
With code interpreterSelf-reported
90.2%
ClawEval-MM
Average score across three trialsSelf-reported
56.9%
CoWorkBench
Self-reported
70.7%
DeepSWE 1.1
Claude Code, temp=1.0, top_p=0.95, 256K contextSelf-reported
42.2%
ERQA
Self-reported
65.5%
Humanity's Last Exam
Judged by GPT-4oSelf-reported
30.8%
IFBench
Self-reported
79.5%
Job Bench
Self-reported
33.4%
LiveCodeBench v6
Self-reported
90.3%
MathVision
With code interpreterSelf-reported
94.6%
NL2Repo
Claude Code harnessSelf-reported
42.3%
OmniDocBench 1.5
Self-reported
91.1%
OSWorld-Verified
Self-reported
84.3%
QwenSWEBench
Claude Code, avg@3, 8-hour timeout, max_tokens=32768, temp=1.0, 256K contextSelf-reported
79.0%
RealWorldQA
Self-reported
85.9%
RecreationBench
Self-reported
47.1%
SWE-Bench Multimodal
Claude Code harness, public SWE-bench Multimodal dev splitSelf-reported
38.6%
SWE-Bench Pro
Claude Code, temp=1.0, top_p=0.95, 256K contextSelf-reported
61.7%
SWE-MM
Claude Code harness on the public SWE-bench Multimodal dev splitSelf-reported
38.6%
Terminal-Bench 2.1
Terminus agent scaffoldSelf-reported
73.0%
Vision2Web
Claude Code harness, judged by gpt-5.4-2026-03-05Self-reported
62.9%
WebArena-Verified
Official WebArena-Verified grader under the OSWorld scaffoldSelf-reported
64.8%

License & Metadata

License
apache-2.0
Announcement Date
August 14, 2026
Last Updated
August 29, 2026

Compare Qwen3.8-27B

All comparisons

Similar Models

All Models

Recommendations are based on similarity of characteristics: developer organization, multimodality, parameter size, and benchmark performance. Choose a model to compare or go to the full catalog to browse all available AI models.