Qwen3 VL 4B Thinking
MultimodalQwen3 VL 4B Thinking is the reasoning variant of the compact Qwen3 VL 4B vision-language model, released in September 2025. It applies extended chain-of-thought to multimodal understanding tasks, scoring 0.94 on DocVQA and 0.87 on MMBench-V1.1. It is Apache 2.0 licensed and open-weight.
Key Specifications
Parameters
4.0B
Context
262.1K
Release Date
September 22, 2025
Average Score
91.3%
Timeline
Key dates in the model's history
Announcement
September 22, 2025
Last Update
August 27, 2026
Technical Specifications
Parameters
4.0B
Training Tokens
-
Knowledge Cutoff
-
Family
-
Capabilities
MultimodalZeroEval
Pricing & Availability
Input (per 1M tokens)
$0.10
Output (per 1M tokens)
$1.00
Max Input Tokens
262.1K
Max Output Tokens
262.1K
Supported Features
Function CallingStructured OutputCode ExecutionWeb SearchBatch InferenceFine-tuning
Benchmark Results
Model performance metrics across various tests and benchmarks
Other Tests
Specialized benchmarks
DocVQA
Self-reported by the model provider • Self-reported
ScreenSpot
Self-reported by the model provider • Self-reported
MMBench-V1.1
en, V1.1 • Self-reported
License & Metadata
License
apache-2.0
Announcement Date
September 22, 2025
Last Updated
August 27, 2026
Compare Qwen3 VL 4B Thinking
All comparisonsSimilar Models
All ModelsQwen3 VL 4B Instruct
Alibaba
MM4.0B
Released:Sep 2025
Price:$0.10/1M tokens
Qwen2.5-Omni-7B
Alibaba
MM7.0B
Best score:0.8 (HumanEval)
Released:Mar 2025
Qwen2.5 VL 7B Instruct
Alibaba
MM8.3B
Released:Jan 2025
Qwen3.5-397B-A17B
Alibaba
MM397.0B
Released:Feb 2026
Qwen3.8 Max
Alibaba
MM2.4T
Released:Aug 2026
Price:$2.50/1M tokens
Qwen2-VL-72B-Instruct
Alibaba
MM73.4B
Released:Aug 2024
Qwen2.5 VL 72B Instruct
Alibaba
MM72.0B
Released:Jan 2025
Qwen2.5 VL 32B Instruct
Alibaba
MM33.5B
Best score:0.9 (HumanEval)
Released:Feb 2025
Recommendations are based on similarity of characteristics: developer organization, multimodality, parameter size, and benchmark performance. Choose a model to compare or go to the full catalog to browse all available AI models.