Key Specifications
Parameters
35.0B
Context
-
Release Date
April 16, 2026
Average Score
68.9%
Timeline
Key dates in the model's history
Announcement
April 16, 2026
Last Update
August 29, 2026
Technical Specifications
Parameters
35.0B
Training Tokens
-
Knowledge Cutoff
-
Family
-
Capabilities
MultimodalZeroEval
Benchmark Results
Model performance metrics across various tests and benchmarks
Programming
Programming skills tests
SWE-Bench Verified
Internal agent scaffold (bash + file-edit), 200K ctx, temp=1.0 • Self-reported
Reasoning
Logical reasoning and analysis
GPQA
• Self-reported
Multimodal
Working with images and visual data
AI2D
TEST • Self-reported
MMMU
• Self-reported
Other Tests
Specialized benchmarks
AIME 2026
Full AIME 2026 (I & II) • Self-reported
CC-OCR
• Self-reported
C-Eval
• Self-reported
CharXiv-R
RQ • Self-reported
Claw-Eval
Pass^3 • Self-reported
DeepPlanning
• Self-reported
EmbSpatialBench
• Self-reported
Hallusion Bench
• Self-reported
HMMT 2025
February 2025 • Self-reported
HMMT25
November 2025 • Self-reported
HMMT Feb 26
• Self-reported
Humanity's Last Exam
• Self-reported
IMO-AnswerBench
• Self-reported
LiveCodeBench v6
• Self-reported
LVBench
• Self-reported
MathVista-Mini
• Self-reported
MCP Atlas
Public set, gemini-2.5-pro judge • Self-reported
MCP-Mark
GitHub MCP v0.30.3, Playwright responses truncated at 32K • Self-reported
MLVU
• Self-reported
MMBench-V1.1
EN_V1.1_dev • Self-reported
MMLU-Pro
• Self-reported
MMLU-Redux
• Self-reported
MMMU-Pro
• Self-reported
MVBench
• Self-reported
NL2Repo
• Self-reported
ODinW
13 • Self-reported
OmniDocBench 1.5
• Self-reported
RealWorldQA
• Self-reported
RefCOCO-avg
• Self-reported
RefSpatialBench
• Self-reported
SimpleVQA
• Self-reported
SkillsBench
OpenCode, 78 self-contained tasks, avg of 5 runs • Self-reported
SuperGPQA
• Self-reported
SWE-bench Multilingual
• Self-reported
SWE-Bench Pro
Refined public set • Self-reported
TAU3-Bench
Official user model (gpt-5.2 low) + BM25 retrieval • Self-reported
Terminal-Bench 2.0
Harbor/Terminus-2, 256K ctx, avg of 5 runs • Self-reported
Toolathlon
• Self-reported
VideoMME w/o sub.
Without subtitles • Self-reported
VideoMME w sub.
With subtitles • Self-reported
VideoMMMU
• Self-reported
VITA-Bench
claude-4-sonnet judge • Self-reported
WideSearch
• Self-reported
ZClawBench
Internal real-user-distribution Claw agent benchmark, 256K ctx • Self-reported
ZEROBench-Sub
• Self-reported
License & Metadata
License
apache-2.0
Announcement Date
April 16, 2026
Last Updated
August 29, 2026
Similar Models
All ModelsQwen3.8-27B
Alibaba
MM27.8B
Best score:0.9 (GPQA)
Released:Aug 2026
Qwen2-VL-72B-Instruct
Alibaba
MM73.4B
Released:Aug 2024
Qwen3 VL 32B Thinking
Alibaba
MM33.0B
Released:Sep 2025
Qwen3.6-27B
Alibaba
MM27.8B
Released:Apr 2026
Price:$0.60/1M tokens
Qwen2.5 VL 72B Instruct
Alibaba
MM72.0B
Released:Jan 2025
Qwen2.5 VL 32B Instruct
Alibaba
MM33.5B
Best score:0.9 (HumanEval)
Released:Feb 2025
QvQ-72B-Preview
Alibaba
MM73.4B
Released:Dec 2024
Qwen3.7-Plus
Alibaba
MM
Best score:0.9 (GPQA)
Released:May 2026
Price:$0.32/1M tokens
Recommendations are based on similarity of characteristics: developer organization, multimodality, parameter size, and benchmark performance. Choose a model to compare or go to the full catalog to browse all available AI models.