llmboard.aiLeaderboard Center
Overall
Overall RankingOpen Models
Tools
Model DirectoryCompare Models
Capabilities
CodingReasoningMathKnowledgeInstruction Following
Price & Efficiency
Price & ValueCapability vs. PriceRuntime Performance
Modalities
Image GenerationVideo GenerationSpeech ModelsEmbeddings
Core Benchmarks
GPQAMMLU-ProAIME 2025SWE-Bench VerifiedMMLUHumanity's Last ExamLiveCodeBenchMATHHumanEvalMMMU-ProView all benchmarks
Methods
Scoring & Data
393 models668 benchmarks

Leaderboard Center

Overall RankingCodingCore BenchmarksPrice & ValueRuntime Performance

Modalities

All ModelsImage GenerationImage EditingVideo GenerationImage-to-VideoVideo EditingText-to-SpeechSpeech-to-TextEmbeddings

Data & Methods

Scoring MethodAll BenchmarksReasoningMath

Vendors

All VendorsOpenAIAnthropicGoogle
llmboard.aiCopyright 2026 llmboard.ai

Alibaba Cloud / Qwen Team model product

Qwen3 VL 32B Thinking

Qwen3-VL is a large multimodal model that unifies vision, language, and reasoning to achieve human-level perception and cognition across text, images, and video.

Updated Aug 12, 2026. Default version: Qwen3 VL 32B Thinking

Compare
LLMBoard score40.5Qwen3 VL 32B Thinking
Coverage100%47 benchmark families
Context windowN/ATokens
Official input priceN/AOfficial price unavailable

On this page

  • Specification
  • Capability
  • Benchmarks
  • Arena
  • Runtime
  • Pricing
  • Versions
  • About
  • Compare
  • Similar models
  • FAQ

Qwen3 VL 32B Thinking Specifications

Technical details for the model's default version.

Version
Qwen3 VL 32B Thinking
Released
Sep 22, 2025
Knowledge cutoff
Unknown
Parameters
33B
Context window
N/A
Max output
N/A
Inputs
image, text
Outputs
text
Open weights
No
License
Apache 2.0

Qwen3 VL 32B Thinking Capability Profile

This profile uses the model's current scored version. Arena ratings and prices are shown separately.

Qwen3 VL 32B Thinking category scores

Qwen3 VL 32B Thinking Benchmark Results

Benchmark scores for Qwen3 VL 32B Thinking.

47 rows
Columns

Show columns

MMMU (val)78.1%0111100.0%CAug 11, 2026
MuirBench80.3%0111100.0%CAug 11, 2026
OCRBench-V2 (en)68.4%0112100.0%CAug 11, 2026
MM-MT-Bench8.3 points021793.8%CAug 11, 2026
OCRBench-V2 (zh)62.1%021190.0%CAug 11, 2026
ScreenSpot95.7%021693.3%CAug 11, 2026
CharadesSTA62.8%041272.7%CAug 11, 2026
CharXiv-D90.2%041680.0%CAug 11, 2026
InfoVQAtest89.2%041272.7%CAug 11, 2026
Multi-IF78.0%042084.2%CAug 11, 2026
WritingBench86.2%041578.6%CAug 11, 2026
AndroidWorld_SR63.7%05842.9%CAug 11, 2026
DocVQAtest96.1%051160.0%CAug 11, 2026
Hallusion Bench67.4%051673.3%CAug 11, 2026
BLINK68.5%061358.3%CAug 11, 2026
MMBench-V1.190.8%061870.6%CAug 11, 2026
MMStar79.4%062276.2%CAug 11, 2026
LiveBench 2024112574.7%071453.9%CAug 11, 2026
MathVista-Mini85.9%072372.7%CAug 11, 2026
MVBench73.2%071762.5%CAug 11, 2026
VideoMME w/o sub.77.3%071033.3%CAug 11, 2026
BFCL-v371.7%081961.1%CAug 11, 2026
Creative Writing v383.3%101325.0%CAug 11, 2026
OSWorld41.0%102052.6%CAug 11, 2026
SimpleQA55.4%104680.0%CAug 11, 2026
Arena-Hard v260.5%111633.3%CAug 11, 2026
PolyMATH52.0%112354.5%CAug 11, 2026
LVBench62.6%132447.8%CAug 11, 2026
ERQA52.3%142340.9%CAug 11, 2026
RealWorldQA78.4%142648.0%CAug 11, 2026
AI2D88.9%153254.8%CAug 11, 2026
Include76.3%153153.3%CAug 11, 2026
MMLU-ProX77.2%153254.8%CAug 11, 2026
OCRBench85.5%152233.3%CAug 11, 2026
MathVision70.2%173248.4%CAug 11, 2026
MMLU88.7%1810082.8%CAug 11, 2026
SuperGPQA59.0%183448.5%CAug 11, 2026
VideoMMMU79.0%182632.0%CAug 11, 2026
ScreenSpot Pro57.1%192525.0%CAug 11, 2026
MMLU-Redux91.9%204859.6%CAug 11, 2026
IFEval87.8%306554.7%CAug 11, 2026
CharXiv-R65.2%344829.8%CAug 11, 2026
LiveCodeBench v665.6%355334.6%CAug 11, 2026
MMLU-Pro82.1%3812971.1%CAug 11, 2026
MMMU-Pro68.1%406640.0%CAug 11, 2026
AIME 202583.7%6311445.1%CAug 11, 2026
GPQA73.1%11923449.4%CAug 11, 2026

Qwen3 VL 32B Thinking Arena Results

Preference and agent-evaluation results for the default version.

No Arena results

The default version does not have a matching Arena result yet.

Qwen3 VL 32B Thinking Runtime Performance

Provider-specific output speed and catalog latency for Qwen3 VL 32B Thinking. Runtime does not affect the capability score.

No runtime data

No provider-specific speed or latency record is linked to the default version yet.

Browse runtime rankings

Output Speed is generated output tokens received per second. Catalog latency is reported separately from observed provider TTFT.

Qwen3 VL 32B Thinking Pricing

Official vendor API pricing appears first, followed by individual provider offers.

Official API
N/A
Official provider
N/A
Lowest third-party
N/A
Tracked offerings
0
No provider prices

The default version has no current input or output token prices.

Official prices use only the vendor's configured official Provider and positive standard USD PAYG rates. Third-party offers remain explicitly labeled.

Qwen3 VL 32B Thinking Versions

Available versions of this model. The score column identifies the version used in the overall ranking.

1 rows
Columns

Show columns

Qwen3 VL 32B ThinkingSep 22, 202540.533BN/AN/ANoApache 2.0

What is Qwen3 VL 32B Thinking?

Key information about Qwen3 VL 32B Thinking and its available data.

Qwen3-VL is a large multimodal model that unifies vision, language, and reasoning to achieve human-level perception and cognition across text, images, and video. Built on a 235B-parameter architecture, it integrates early joint training of visual and textual modalities for strong language grounding.

The model supports up to a 1 million-token context window and excels at visual understanding, spatial reasoning, long video comprehension, and tool-based interaction. It can generate code from images, perform precise 2D/3D object grounding, and operate digital interfaces like a visual agent.

5 Pro in perception benchmarks, while the “Thinking” version leads in multimodal reasoning and STEM tasks. With multilingual OCR, creative writing, and fine-grained scene interpretation, Qwen3-VL establishes a new open-source frontier for integrated vision-language intelligence.

Data as of 2026-08-11.

Qwen3 VL 32B Thinking vs nearby models

Open a comparison with the three ranked models immediately above and below this model.

Qwen3 VL 32B ThinkingvsQwen3 Next 80B A3B ThinkingQwen3 VL 32B ThinkingvsGrok 4.3Qwen3 VL 32B ThinkingvsQwen3 MaxQwen3 VL 32B ThinkingvsDeepSeek-V3.1Qwen3 VL 32B ThinkingvsClaude Sonnet 4Qwen3 VL 32B ThinkingvsMercury 2

Models similar to Qwen3 VL 32B Thinking

Recommendations prioritize the same model type and family, then the closest LLMBoard score.

#126+0.2
AC

Qwen3 Max

Alibaba Cloud / Qwen Team

40.7 LLMBoard

DetailsCompare
#124+0.4
AC

Qwen3 Next 80B A3B Thinking

Alibaba Cloud / Qwen Team

40.9 LLMBoard

DetailsCompare
#123+0.9
AC

Qwen3.5 9B

Alibaba Cloud / Qwen Team

41.4 LLMBoard

DetailsCompare
#119+1.9
AC

Qwen3 235B A22B

Alibaba Cloud / Qwen Team

42.4 LLMBoard

DetailsCompare
#117+2.7
AC

Qwen3 VL 235B A22B

Alibaba Cloud / Qwen Team

43.2 LLMBoard

DetailsCompare
#139-3.9
AC

Qwen3 Next 80B A3B

Alibaba Cloud / Qwen Team

36.6 LLMBoard

DetailsCompare

FAQ

Common questions about Qwen3 VL 32B Thinking.

When was Qwen3 VL 32B Thinking released?

Qwen3 VL 32B Thinking's default version was released on Sep 22, 2025.

How much does Qwen3 VL 32B Thinking cost?

No official standard PAYG price is currently available for Qwen3 VL 32B Thinking.

Who created Qwen3 VL 32B Thinking?

Qwen3 VL 32B Thinking was created by Alibaba Cloud / Qwen Team.

What is the context window for Qwen3 VL 32B Thinking?

A context window is not available for the default version.

Is Qwen3 VL 32B Thinking open weight?

No. The default version is not marked as having publicly available weights.

How many API providers offer Qwen3 VL 32B Thinking?

No provider offering is currently linked to the default version.

What models should I compare Qwen3 VL 32B Thinking with?

Nearby ranked alternatives include Qwen3 Next 80B A3B Thinking, Grok 4.3, Qwen3 Max.