llmboard.aiAI model intelligence
Home

Model Rankings

OverallOpen ModelsAgentCodingReasoningMathKnowledgeInstruction FollowingTextVision
Image GenerationImage Editing
Video GenerationImage to VideoVideo Editing
Text to SpeechSpeech to Text
Embeddings

Efficiency

Chat Token PricingImage PricingVideo PricingAudio Pricing
Chat Speed & LatencyProvider Reliability

Benchmarks

GPQAMMLU-ProAIME 2025SWE-Bench VerifiedMMLUHumanity's Last ExamLiveCodeBenchMATHHumanEvalMMMU-Pro
All Benchmarks

Tools

Model DirectoryCompare Models

Scoring & Data

Scoring & Data
393 models668 benchmarks

Leaderboard Center

Overall RankingCodingCore BenchmarksPrice & ValueRuntime Performance

Modalities

All ModelsImage GenerationImage EditingVideo GenerationImage-to-VideoVideo EditingText-to-SpeechSpeech-to-TextEmbeddings

Data & Methods

Scoring MethodAll BenchmarksReasoningMath

Vendors

All VendorsOpenAIAnthropicGoogle
llmboard.aiCopyright 2026 llmboard.ai

Alibaba Cloud / Qwen Team model product

Qwen3 32B

Qwen3-32B is a large language model from Alibaba's Qwen3 series.

Updated Aug 17, 2026. Default version: Qwen3 32B

Compare
LLMBoard Score30.8Qwen3 32B
Coverage40%8 benchmark families
Context window131.1KTokens
Official input price$0.70Alibaba API

On this page

  • Capability
  • Benchmarks
  • Arena
  • Pricing
  • Runtime
  • Specification
  • Versions
  • Compare
  • Similar models
  • About
  • FAQ

Qwen3 32B Capability Profile

This profile uses the model's current scored version. Arena ratings and prices are shown separately.

Qwen3 32B LLMBoard score breakdown

Qwen3 32B Benchmark Results

Benchmark scores for Qwen3 32B.

9 rows
Columns

Show columns

Sort by
Benchmark
Score
Rank
Participants
Percentile
Evidence
Evaluated
BenchmarkMultiLFScore73.0%Rank01Participants2Percentile100.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkArena HardScore93.8%Rank02Participants26Percentile96.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkAiderScore50.2%Rank04Participants4Percentile0.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkBFCLScore70.3%Rank06Participants11Percentile50.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkCodeForcesScore65.9%Rank14Participants17Percentile18.8%EvidenceCEvaluatedAug 17, 2026
BenchmarkLiveBenchScore74.9%Rank15Participants38Percentile62.2%EvidenceCEvaluatedAug 17, 2026
BenchmarkAIME 2024Score81.4%Rank23Participants53Percentile57.7%EvidenceCEvaluatedAug 17, 2026
BenchmarkLiveCodeBenchScore65.7%Rank28Participants75Percentile63.5%EvidenceCEvaluatedAug 17, 2026
BenchmarkAIME 2025Score72.9%Rank82Participants115Percentile28.9%EvidenceCEvaluatedAug 17, 2026

Qwen3 32B Arena Results

Preference and agent-evaluation results for the default version.

2 rows
Columns

Show columns

Sort by
Arena
Category
Rank
Rating / score
Votes
Observations
Result date
ArenatextCategoryoverallRank189Rating / score1340.1Votes3,926ObservationsN/AResult dateAug 12, 2026
Arenatext style controlCategoryoverallRank200Rating / score1347.1Votes3,926ObservationsN/AResult dateAug 12, 2026

Qwen3 32B Pricing

Official vendor API pricing appears first, followed by individual provider offers.

Official API
$0.70 input, $2.8 output per 1M
Official provider
Alibaba
Lowest third-party
From $0.08 input, $0.28 output per 1M via Deep Infra
Tracked offerings
21
20 rows
Columns

Show columns

Sort by
Provider
Provider model ID
Region
Input / 1M
Output / 1M
Context
Updated
ProvideriFlowProvider model IDqwen3-32bRegionglobalInput / 1MN/AOutput / 1MN/AContext128KUpdatedAug 17, 2026
ProviderDeep InfraProvider model IDQwen/Qwen3-32BRegionglobalInput / 1M$0.08Output / 1M$0.28Context41KUpdatedAug 17, 2026
ProviderOpenRouterProvider model IDqwen/qwen3-32bRegionglobalInput / 1M$0.08Output / 1M$0.28Context131.1KUpdatedAug 17, 2026
ProviderKilo GatewayProvider model IDqwen/qwen3-32bRegionglobalInput / 1M$0.08Output / 1M$0.28Context41KUpdatedAug 17, 2026
ProviderOVHcloud AI EndpointsProvider model IDqwen3-32bRegionglobalInput / 1M$0.09Output / 1M$0.25Context32.8KUpdatedAug 17, 2026
ProviderAbacusProvider model IDQwen/Qwen3-32BRegionglobalInput / 1M$0.09Output / 1M$0.29Context131.1KUpdatedAug 17, 2026
ProviderCortecsProvider model IDqwen3-32bRegionglobalInput / 1M$0.099Output / 1M$0.299Context40KUpdatedAug 17, 2026
ProviderNanoGPTProvider model IDqwen/qwen3-32bRegionglobalInput / 1M$0.10Output / 1M$0.30Context41KUpdatedAug 17, 2026
ProviderJiekou.AIProvider model IDqwen/qwen3-32b-fp8RegionglobalInput / 1M$0.10Output / 1M$0.45Context41KUpdatedAug 17, 2026
ProviderNebius Token FactoryProvider model IDQwen/Qwen3-32BRegionglobalInput / 1M$0.10Output / 1M$0.30Context128KUpdatedAug 17, 2026
ProviderNovitaAIProvider model IDqwen/qwen3-32b-fp8RegionglobalInput / 1M$0.10Output / 1M$0.45Context41KUpdatedAug 17, 2026
ProviderLLM GatewayProvider model IDqwen3-32bRegionglobalInput / 1M$0.10Output / 1M$0.30Context41KUpdatedAug 17, 2026
ProviderMerge GatewayProvider model IDqwen/qwen3-32bRegionglobalInput / 1M$0.15Output / 1M$0.60Context131.1KUpdatedAug 17, 2026
ProviderVercel AI GatewayProvider model IDalibaba/qwen-3-32bRegionglobalInput / 1M$0.16Output / 1M$0.64Context128KUpdatedAug 17, 2026
ProviderDigitalOceanProvider model IDalibaba-qwen3-32bRegionglobalInput / 1M$0.25Output / 1M$0.55Context32.8KUpdatedAug 17, 2026
ProviderAlibaba (China)Provider model IDqwen3-32bRegionglobalInput / 1M$0.287Output / 1M$1.15Context131.1KUpdatedAug 17, 2026
ProviderHugging FaceProvider model IDQwen/Qwen3-32BRegionglobalInput / 1M$0.29Output / 1M$0.59Context131.1KUpdatedAug 17, 2026
ProviderHeliconeProvider model IDqwen3-32bRegionglobalInput / 1M$0.29Output / 1M$0.59Context131.1KUpdatedAug 17, 2026
ProviderAlibabaProvider model IDqwen3-32bRegionglobalInput / 1M$0.70Output / 1M$2.8Context131.1KUpdatedAug 17, 2026
ProviderPioneerProvider model IDQwen/Qwen3-32BRegionglobalInput / 1M$0.90Output / 1M$0.90Context131.1KUpdatedAug 17, 2026

Official prices use only the vendor's configured official Provider and positive standard USD PAYG rates. Third-party offers remain explicitly labeled.

Qwen3 32B Runtime Performance

Provider-specific output speed and catalog latency for Qwen3 32B. Runtime does not affect the capability score.

3 rows
Columns

Show columns

Sort by
Provider
Output Speed
Catalog Latency
Max Input
Max Output
Updated
ProviderSambanovaOutput Speed327.7 tok/sCatalog Latency1.08 sMax Input128KMax Output128KUpdatedAug 17, 2026
ProviderNovitaOutput Speed32.43 tok/sCatalog Latency0.93 sMax Input128KMax Output128KUpdatedAug 17, 2026
ProviderDeepInfraOutput Speed26.95 tok/sCatalog Latency1.19 sMax Input128KMax Output128KUpdatedAug 17, 2026

Output Speed is generated output tokens received per second. Catalog latency is reported separately from observed provider TTFT.

Qwen3 32B Specifications

Technical details for the model's default version.

Version
Qwen3 32B
Released
Apr 29, 2025
Knowledge cutoff
Apr 1, 2025
Parameters
32.8B
Context window
131.1K
Max output
16.4K
Inputs
text
Outputs
text
Open weights
Yes
License
Apache 2.0

Qwen3 32B Versions

Available versions of this model. The score column identifies the version used in the overall ranking.

1 rows
Columns

Show columns

Sort by
Version
Released
LLMBoard
Parameters
Context
Max output
Open weights
License
VersionQwen3 32BReleasedApr 29, 2025LLMBoard30.8Parameters32.8BContext131.1KMax output16.4KOpen weightsYesLicenseApache 2.0

Qwen3 32B vs nearby models

Open a comparison with the three ranked models immediately above and below this model.

Qwen3 32BvsClaude Sonnet 3.5Qwen3 32BvsGPT-4.1Qwen3 32BvsQwen3 VL 30B A3B ThinkingQwen3 32BvsGPT-5-nanoQwen3 32BvsLlama 3.1 Nemotron Ultra 253BQwen3 32BvsKimi k1.5

Models similar to Qwen3 32B

Recommendations prioritize the same model type and family, then the closest LLMBoard score.

#169+0.3
AC

Qwen3 VL 30B A3B Thinking

Alibaba Cloud / Qwen Team

31.1 LLMBoard

DetailsCompare
#165+1.2
AC

Qwen3.5 4B

Alibaba Cloud / Qwen Team

32.0 LLMBoard

DetailsCompare
#176-2.6
AC

Qwen3 30B A3B

Alibaba Cloud / Qwen Team

28.2 LLMBoard

DetailsCompare
#177-2.7
AC

Qwen3 VL 30B A3B

Alibaba Cloud / Qwen Team

28.1 LLMBoard

DetailsCompare
#178-3.0
AC

Qwen3 VL 8B Thinking

Alibaba Cloud / Qwen Team

27.8 LLMBoard

DetailsCompare
#158+3.8
AC

Qwen3 VL 32B

Alibaba Cloud / Qwen Team

34.6 LLMBoard

DetailsCompare

What is Qwen3 32B?

Key information about Qwen3 32B and its available data.

Qwen3-32B is a large language model from Alibaba's Qwen3 series. 8 billion parameters, a 128k token context window, support for 119 languages, and hybrid thinking modes allowing switching between deep reasoning and fast responses.

It demonstrates strong performance in reasoning, instruction-following, and agent capabilities.

Data as of 2026-08-17.

FAQ

Common questions about Qwen3 32B.

When was Qwen3 32B released?

Qwen3 32B's default version was released on Apr 29, 2025.

How much does Qwen3 32B cost?

Qwen3 32B's official API price is $0.70 per million input tokens and $2.8 per million output tokens via Alibaba. The lowest tracked third-party offer starts at $0.08 input and $0.28 output via Deep Infra.

Who created Qwen3 32B?

Qwen3 32B was created by Alibaba Cloud / Qwen Team.

What is the context window for Qwen3 32B?

The default version has a 131.1K token context window.

Is Qwen3 32B open weight?

Yes. The default version is marked as open weight under Apache 2.0.

How many API providers offer Qwen3 32B?

21 provider offerings are linked to the default version.

What models should I compare Qwen3 32B with?

Nearby ranked alternatives include Claude Sonnet 3.5, GPT-4.1, Qwen3 VL 30B A3B Thinking.