llmboard.aiAI model intelligence
Home

Model Rankings

OverallOpen ModelsAgentCodingReasoningMathKnowledgeInstruction FollowingTextVision
Image GenerationImage Editing
Video GenerationImage to VideoVideo Editing
Text to SpeechSpeech to Text
Embeddings

Efficiency

Chat Token PricingImage PricingVideo PricingAudio Pricing
Chat Speed & LatencyProvider Reliability

Benchmarks

GPQAMMLU-ProAIME 2025SWE-Bench VerifiedMMLUHumanity's Last ExamLiveCodeBenchMATHHumanEvalMMMU-Pro
All Benchmarks

Tools

Model DirectoryCompare Models

Scoring & Data

Scoring & Data
393 models668 benchmarks

Leaderboard Center

Overall RankingCodingCore BenchmarksPrice & ValueRuntime Performance

Modalities

All ModelsImage GenerationImage EditingVideo GenerationImage-to-VideoVideo EditingText-to-SpeechSpeech-to-TextEmbeddings

Data & Methods

Scoring MethodAll BenchmarksReasoningMath

Vendors

All VendorsOpenAIAnthropicGoogle
llmboard.aiCopyright 2026 llmboard.ai

Alibaba Cloud / Qwen Team model product

Qwen3.5 4B

5-4B is a 4 billion parameter vision-language model using Gated DeltaNet hybrid architecture with a 3:1 ratio of linear attention to full softmax attention.

Updated Aug 17, 2026. Default version: Qwen3.5-4B

Compare
LLMBoard Score32.0Qwen3.5-4B
Coverage100%25 benchmark families
Context windowN/ATokens
Official input priceN/AOfficial price unavailable

On this page

  • Capability
  • Benchmarks
  • Arena
  • Pricing
  • Runtime
  • Specification
  • Versions
  • Compare
  • Similar models
  • About
  • FAQ

Qwen3.5 4B Capability Profile

This profile uses the model's current scored version. Arena ratings and prices are shown separately.

Qwen3.5-4B LLMBoard score breakdown

Qwen3.5 4B Benchmark Results

Benchmark scores for Qwen3.5-4B.

25 rows
Columns

Show columns

Sort by
Benchmark
Score
Rank
Participants
Percentile
Evidence
Evaluated
BenchmarkDeepPlanningScore17.6%Rank09Participants9Percentile0.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkMAXIFEScore78.0%Rank09Participants11Percentile20.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkNOVA-63Score54.3%Rank09Participants11Percentile20.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkVITA-BenchScore22.0%Rank10Participants10Percentile0.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkGlobal PIQAScore78.9%Rank11Participants13Percentile16.7%EvidenceCEvaluatedAug 17, 2026
BenchmarkBFCL-V4Score50.3%Rank12Participants15Percentile21.4%EvidenceCEvaluatedAug 17, 2026
Benchmarkt2-benchScore79.9%Rank12Participants23Percentile50.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkC-EvalScore85.1%Rank13Participants18Percentile29.4%EvidenceCEvaluatedAug 17, 2026
BenchmarkPolyMATHScore51.1%Rank13Participants23Percentile45.5%EvidenceCEvaluatedAug 17, 2026
BenchmarkWMT24++Score66.6%Rank13Participants23Percentile45.5%EvidenceCEvaluatedAug 17, 2026
BenchmarkHMMT25Score76.8%Rank14Participants25Percentile45.8%EvidenceCEvaluatedAug 17, 2026
BenchmarkLongBench v2Score50.0%Rank14Participants17Percentile18.8%EvidenceCEvaluatedAug 17, 2026
BenchmarkAA-LCRScore57.0%Rank15Participants18Percentile17.6%EvidenceCEvaluatedAug 17, 2026
BenchmarkMulti-ChallengeScore49.0%Rank17Participants29Percentile42.9%EvidenceCEvaluatedAug 17, 2026
BenchmarkIFEvalScore89.8%Rank18Participants67Percentile74.2%EvidenceCEvaluatedAug 17, 2026
BenchmarkMMLU-ProXScore71.5%Rank20Participants32Percentile38.7%EvidenceCEvaluatedAug 17, 2026
BenchmarkIncludeScore71.0%Rank21Participants31Percentile33.3%EvidenceCEvaluatedAug 17, 2026
BenchmarkSuperGPQAScore52.9%Rank26Participants34Percentile24.2%EvidenceCEvaluatedAug 17, 2026
BenchmarkHMMT 2025Score74.0%Rank27Participants33Percentile18.8%EvidenceCEvaluatedAug 17, 2026
BenchmarkIFBenchScore59.2%Rank28Participants34Percentile18.2%EvidenceCEvaluatedAug 17, 2026
BenchmarkMMLU-ReduxScore88.8%Rank29Participants48Percentile40.4%EvidenceCEvaluatedAug 17, 2026
BenchmarkMMMLUScore76.1%Rank42Participants49Percentile14.6%EvidenceCEvaluatedAug 17, 2026
BenchmarkLiveCodeBench v6Score55.8%Rank43Participants56Percentile23.6%EvidenceCEvaluatedAug 17, 2026
BenchmarkMMLU-ProScore79.1%Rank60Participants134Percentile55.6%EvidenceCEvaluatedAug 17, 2026
BenchmarkGPQAScore76.2%Rank104Participants239Percentile56.7%EvidenceCEvaluatedAug 17, 2026

Qwen3.5 4B Arena Results

Preference and agent-evaluation results for the default version.

No Arena results

The default version does not have a matching Arena result yet.

Qwen3.5 4B Pricing

Official vendor API pricing appears first, followed by individual provider offers.

Official API
N/A
Official provider
N/A
Lowest third-party
From $0.04 input, $0.07 output per 1M via EmpirioLabs AI
Tracked offerings
3
3 rows
Columns

Show columns

Sort by
Provider
Provider model ID
Region
Input / 1M
Output / 1M
Context
Updated
ProviderQVACProvider model IDqwen3.5-4bRegionglobalInput / 1MN/AOutput / 1MN/AContext32.8KUpdatedAug 17, 2026
ProviderEmpirioLabs AIProvider model IDqwen3-5-4bRegionglobalInput / 1M$0.04Output / 1M$0.07Context262.1KUpdatedAug 17, 2026
ProviderNanoGPTProvider model IDqwen3.5-4bRegionglobalInput / 1M$0.10Output / 1M$0.20Context262.1KUpdatedAug 17, 2026

Official prices use only the vendor's configured official Provider and positive standard USD PAYG rates. Third-party offers remain explicitly labeled.

Qwen3.5 4B Runtime Performance

Provider-specific output speed and catalog latency for Qwen3.5-4B. Runtime does not affect the capability score.

No runtime data

No provider-specific speed or latency record is linked to the default version yet.

Browse runtime rankings

Output Speed is generated output tokens received per second. Catalog latency is reported separately from observed provider TTFT.

Qwen3.5 4B Specifications

Technical details for the model's default version.

Version
Qwen3.5-4B
Released
Mar 2, 2026
Knowledge cutoff
Unknown
Parameters
4B
Context window
N/A
Max output
N/A
Inputs
image, text
Outputs
text
Open weights
No
License
Apache 2.0

Qwen3.5 4B Versions

Available versions of this model. The score column identifies the version used in the overall ranking.

1 rows
Columns

Show columns

Sort by
Version
Released
LLMBoard
Parameters
Context
Max output
Open weights
License
VersionQwen3.5-4BReleasedMar 2, 2026LLMBoard32.0Parameters4BContextN/AMax outputN/AOpen weightsNoLicenseApache 2.0

Qwen3.5 4B vs nearby models

Open a comparison with the three ranked models immediately above and below this model.

Qwen3.5 4BvsLFM2.5 2.6BQwen3.5 4BvsMistral Small 4Qwen3.5 4BvsLongCat Flash LiteQwen3.5 4BvsNemotron Nano 9BQwen3.5 4BvsClaude Sonnet 3.5Qwen3.5 4BvsGPT-4.1

Models similar to Qwen3.5 4B

Recommendations prioritize the same model type and family, then the closest LLMBoard score.

#169-0.9
AC

Qwen3 VL 30B A3B Thinking

Alibaba Cloud / Qwen Team

31.1 LLMBoard

DetailsCompare
#170-1.2
AC

Qwen3 32B

Alibaba Cloud / Qwen Team

30.8 LLMBoard

DetailsCompare
#158+2.6
AC

Qwen3 VL 32B

Alibaba Cloud / Qwen Team

34.6 LLMBoard

DetailsCompare
#155+3.1
AC

Qwen3 Coder 480B A35B

Alibaba Cloud / Qwen Team

35.0 LLMBoard

DetailsCompare
#176-3.8
AC

Qwen3 30B A3B

Alibaba Cloud / Qwen Team

28.2 LLMBoard

DetailsCompare
#177-3.9
AC

Qwen3 VL 30B A3B

Alibaba Cloud / Qwen Team

28.1 LLMBoard

DetailsCompare

What is Qwen3.5 4B?

Key information about Qwen3.5 4B and its available data.

5-4B is a 4 billion parameter vision-language model using Gated DeltaNet hybrid architecture with a 3:1 ratio of linear attention to full softmax attention. It supports 262K native context length and delivers strong performance for its size across knowledge, reasoning, coding, and multilingual tasks.

Data as of 2026-08-17.

FAQ

Common questions about Qwen3.5 4B.

When was Qwen3.5 4B released?

Qwen3.5 4B's default version was released on Mar 2, 2026.

How much does Qwen3.5 4B cost?

No official standard PAYG price is currently available for Qwen3.5 4B. The lowest tracked third-party offer starts at $0.04 input and $0.07 output via EmpirioLabs AI.

Who created Qwen3.5 4B?

Qwen3.5 4B was created by Alibaba Cloud / Qwen Team.

What is the context window for Qwen3.5 4B?

A context window is not available for the default version.

Is Qwen3.5 4B open weight?

No. The default version is not marked as having publicly available weights.

How many API providers offer Qwen3.5 4B?

3 provider offerings are linked to the default version.

What models should I compare Qwen3.5 4B with?

Nearby ranked alternatives include LFM2.5 2.6B, Mistral Small 4, LongCat Flash Lite.