llmboard.aiLeaderboard Center
Overall
Overall RankingOpen Models
Tools
Model DirectoryCompare Models
Capabilities
CodingReasoningMathKnowledgeInstruction Following
Price & Efficiency
Price & ValueCapability vs. PriceRuntime Performance
Modalities
Image GenerationVideo GenerationSpeech ModelsEmbeddings
Core Benchmarks
GPQAMMLU-ProAIME 2025SWE-Bench VerifiedMMLUHumanity's Last ExamLiveCodeBenchMATHHumanEvalMMMU-ProView all benchmarks
Methods
Scoring & Data
393 models668 benchmarks

Leaderboard Center

Overall RankingCodingCore BenchmarksPrice & ValueRuntime Performance

Modalities

All ModelsImage GenerationImage EditingVideo GenerationImage-to-VideoVideo EditingText-to-SpeechSpeech-to-TextEmbeddings

Data & Methods

Scoring MethodAll BenchmarksReasoningMath

Vendors

All VendorsOpenAIAnthropicGoogle
llmboard.aiCopyright 2026 llmboard.ai

Alibaba Cloud / Qwen Team model product

Qwen3 235B A22B Thinking

Qwen3-235B-A22B-Thinking-2507 is a state-of-the-art thinking-enabled Mixture-of-Experts (MoE) model with 235B total parameters (22B activated).

Updated Aug 12, 2026. Default version: Qwen3-235B-A22B-Thinking-2507

Compare
LLMBoard score48.5Qwen3-235B-A22B-Thinking-2507
Coverage100%24 benchmark families
Context window262.1KTokens
Official input priceN/AOfficial price unavailable

On this page

  • Specification
  • Capability
  • Benchmarks
  • Arena
  • Runtime
  • Pricing
  • Versions
  • About
  • Compare
  • Similar models
  • FAQ

Qwen3 235B A22B Thinking Specifications

Technical details for the model's default version.

Version
Qwen3-235B-A22B-Thinking-2507
Released
Jul 25, 2025
Knowledge cutoff
Unknown
Parameters
235B
Context window
262.1K
Max output
131.1K
Inputs
text
Outputs
text
Open weights
No
License
Apache 2.0

Qwen3 235B A22B Thinking Capability Profile

This profile uses the model's current scored version. Arena ratings and prices are shown separately.

Qwen3-235B-A22B-Thinking-2507 category scores

Qwen3 235B A22B Thinking Benchmark Results

Benchmark scores for Qwen3-235B-A22B-Thinking-2507.

25 rows
Columns

Show columns

CFEval2,134 points012100.0%CAug 11, 2026
Multi-IF80.6%0120100.0%CAug 11, 2026
WritingBench88.3%0115100.0%CAug 11, 2026
LiveBench 2024112578.4%021492.3%CAug 11, 2026
Arena-Hard v279.7%031686.7%CAug 11, 2026
Creative Writing v386.1%031383.3%CAug 11, 2026
BFCL-v371.9%061972.2%CAug 11, 2026
OJBench32.5%06937.5%CAug 11, 2026
MMLU-Redux93.8%074887.2%CAug 11, 2026
Include81.0%083176.7%CAug 11, 2026
MMLU-ProX81.0%083277.4%CAug 11, 2026
PolyMATH60.1%082368.2%CAug 11, 2026
HMMT2583.9%112558.3%CAug 11, 2026
SuperGPQA64.9%113469.7%CAug 11, 2026
Tau2 Airline58.0%152336.4%CAug 11, 2026
Tau2 Retail71.9%162640.0%CAug 11, 2026
TAU-bench Airline46.0%172327.3%CAug 11, 2026
TAU-bench Retail67.8%172533.3%CAug 11, 2026
LiveCodeBench v674.1%255353.9%CAug 11, 2026
MMLU-Pro84.4%2512981.3%CAug 11, 2026
IFEval87.8%286557.8%CAug 11, 2026
Tau2 Telecom45.6%313511.8%CAug 11, 2026
AIME 202592.3%3711468.1%CAug 11, 2026
Humanity's Last Exam18.2%619334.8%CAug 11, 2026
GPQA81.1%8023466.1%CAug 11, 2026

Qwen3 235B A22B Thinking Arena Results

Preference and agent-evaluation results for the default version.

2 rows
Columns

Show columns

textoverall1131414.09,018N/AAug 10, 2026
text style controloverall1391399.29,018N/AAug 10, 2026

Qwen3 235B A22B Thinking Runtime Performance

Provider-specific output speed and catalog latency for Qwen3-235B-A22B-Thinking-2507. Runtime does not affect the capability score.

No runtime data

No provider-specific speed or latency record is linked to the default version yet.

Browse runtime rankings

Output Speed is generated output tokens received per second. Catalog latency is reported separately from observed provider TTFT.

Qwen3 235B A22B Thinking Pricing

Official vendor API pricing appears first, followed by individual provider offers.

Official API
N/A
Official provider
N/A
Lowest third-party
From $0.20 input, $0.60 output per 1M via submodel
Tracked offerings
11
10 rows
Columns

Show columns

ModelScopeQwen/Qwen3-235B-A22B-Thinking-2507globalN/AN/A262.1KAug 11, 2026
iFlowqwen3-235b-a22b-thinking-2507globalN/AN/A256KAug 11, 2026
submodelQwen/Qwen3-235B-A22B-Thinking-2507global$0.20$0.60262.1KAug 11, 2026
OpenRouterqwen/qwen3-235b-a22b-thinking-2507global$0.23$2.3262.1KAug 11, 2026
LLM Gatewayqwen3-235b-a22b-thinking-2507global$0.30$3262KAug 11, 2026
NovitaAIqwen/qwen3-235b-a22b-thinking-2507global$0.30$3131.1KAug 11, 2026
Hugging FaceQwen/Qwen3-235B-A22B-Thinking-2507global$0.30$3262.1KAug 11, 2026
Jiekou.AIqwen/qwen3-235b-a22b-thinking-2507global$0.30$3131.1KAug 11, 2026
Vercel AI Gatewayalibaba/qwen3-235b-a22b-thinkingglobal$0.40$4131.1KAug 11, 2026
Venice AIqwen3-235b-a22b-thinking-2507global$0.45$3.5128KAug 11, 2026

Official prices use only the vendor's configured official Provider and positive standard USD PAYG rates. Third-party offers remain explicitly labeled.

Qwen3 235B A22B Thinking Versions

Available versions of this model. The score column identifies the version used in the overall ranking.

1 rows
Columns

Show columns

Qwen3-235B-A22B-Thinking-2507Jul 25, 202548.5235B262.1K131.1KNoApache 2.0

What is Qwen3 235B A22B Thinking?

Key information about Qwen3 235B A22B Thinking and its available data.

Qwen3-235B-A22B-Thinking-2507 is a state-of-the-art thinking-enabled Mixture-of-Experts (MoE) model with 235B total parameters (22B activated). It features 94 layers, 128 experts (8 activated), and supports 262K native context length.

This version delivers significantly improved reasoning performance, achieving state-of-the-art results among open-source thinking models on logical reasoning, mathematics, science, coding, and academic benchmarks.

Key enhancements include markedly better general capabilities (instruction following, tool usage, text generation), enhanced 256K long-context understanding, and increased thinking depth. The model supports only thinking mode with automatic <think> tag inclusion.

Data as of 2026-08-11.

Qwen3 235B A22B Thinking vs nearby models

Open a comparison with the three ranked models immediately above and below this model.

Qwen3 235B A22B ThinkingvsGLM 4.5Qwen3 235B A22B ThinkingvsMistral Medium 3.5Qwen3 235B A22B ThinkingvsGPT-5.4-nanoQwen3 235B A22B ThinkingvsGrok 3 MiniQwen3 235B A22B ThinkingvsMiMo V2.5 ProQwen3 235B A22B ThinkingvsGrok 4.1

Models similar to Qwen3 235B A22B Thinking

Recommendations prioritize the same model type and family, then the closest LLMBoard score.

#104-2.8
AC

Qwen3 VL 235B A22B Thinking

Alibaba Cloud / Qwen Team

45.8 LLMBoard

DetailsCompare
#87+3.5
AC

Qwen3.5 35B A3B

Alibaba Cloud / Qwen Team

52.0 LLMBoard

DetailsCompare
#117-5.4
AC

Qwen3 VL 235B A22B

Alibaba Cloud / Qwen Team

43.2 LLMBoard

DetailsCompare
#119-6.1
AC

Qwen3 235B A22B

Alibaba Cloud / Qwen Team

42.4 LLMBoard

DetailsCompare
#74+6.8
AC

Qwen3.6 35B A3B

Alibaba Cloud / Qwen Team

55.3 LLMBoard

DetailsCompare
#123-7.1
AC

Qwen3.5 9B

Alibaba Cloud / Qwen Team

41.4 LLMBoard

DetailsCompare

FAQ

Common questions about Qwen3 235B A22B Thinking.

When was Qwen3 235B A22B Thinking released?

Qwen3 235B A22B Thinking's default version was released on Jul 25, 2025.

How much does Qwen3 235B A22B Thinking cost?

No official standard PAYG price is currently available for Qwen3 235B A22B Thinking. The lowest tracked third-party offer starts at $0.20 input and $0.60 output via submodel.

Who created Qwen3 235B A22B Thinking?

Qwen3 235B A22B Thinking was created by Alibaba Cloud / Qwen Team.

What is the context window for Qwen3 235B A22B Thinking?

The default version has a 262.1K token context window.

Is Qwen3 235B A22B Thinking open weight?

No. The default version is not marked as having publicly available weights.

How many API providers offer Qwen3 235B A22B Thinking?

11 provider offerings are linked to the default version.

What models should I compare Qwen3 235B A22B Thinking with?

Nearby ranked alternatives include GLM 4.5, Mistral Medium 3.5, GPT-5.4-nano.