llmboard.aiAI model intelligence
Home

Model Rankings

OverallOpen ModelsAgentCodingReasoningMathKnowledgeInstruction FollowingTextVision
Image GenerationImage Editing
Video GenerationImage to VideoVideo Editing
Text to SpeechSpeech to Text
Embeddings

Efficiency

Chat Token PricingImage PricingVideo PricingAudio Pricing
Chat Speed & LatencyProvider Reliability

Benchmarks

GPQAMMLU-ProAIME 2025SWE-Bench VerifiedMMLUHumanity's Last ExamLiveCodeBenchMATHHumanEvalMMMU-Pro
All Benchmarks

Tools

Model DirectoryCompare Models

Scoring & Data

Scoring & Data
393 models668 benchmarks

Leaderboard Center

Overall RankingCodingCore BenchmarksPrice & ValueRuntime Performance

Modalities

All ModelsImage GenerationImage EditingVideo GenerationImage-to-VideoVideo EditingText-to-SpeechSpeech-to-TextEmbeddings

Data & Methods

Scoring MethodAll BenchmarksReasoningMath

Vendors

All VendorsOpenAIAnthropicGoogle
llmboard.aiCopyright 2026 llmboard.ai

Alibaba Cloud / Qwen Team model product

Qwen3.7 Plus

7-Plus is Alibaba Cloud Qwen Team's multimodal agent model that unifies vision and language into a single agent foundation.

Updated Aug 17, 2026. Default version: Qwen3.7-Plus

Compare
LLMBoard Score73.2Qwen3.7-Plus
Coverage100%59 benchmark families
Context window1MTokens
Official input price$0.50Alibaba API

On this page

  • Capability
  • Benchmarks
  • Arena
  • Pricing
  • Runtime
  • Specification
  • Versions
  • Compare
  • Similar models
  • About
  • FAQ

Qwen3.7 Plus Capability Profile

This profile uses the model's current scored version. Arena ratings and prices are shown separately.

Qwen3.7-Plus LLMBoard score breakdown

Qwen3.7 Plus Benchmark Results

Benchmark scores for Qwen3.7-Plus.

30 of 70 rows
Columns

Show columns

Sort by
Benchmark
Score
Rank
Participants
Percentile
Evidence
Evaluated
BenchmarkBC-VLScore51.1%Rank01Participants1Percentile100.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkCountQAScore77.0%Rank01Participants1Percentile100.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkDeepPlanningScore62.3%Rank01Participants9Percentile100.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkHiPhOScore84.1%Rank01Participants1Percentile100.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkLingoQAScore83.4%Rank01Participants4Percentile100.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkMedXpertQA-MMScore71.0%Rank01Participants1Percentile100.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkMLVUScore87.4%Rank01Participants10Percentile100.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkMMBCScore46.3%Rank01Participants1Percentile100.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkMMSearch-PlusScore41.4%Rank01Participants2Percentile100.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkMRCR v2Score91.7%Rank01Participants3Percentile100.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkOCRBench_V2Score67.1%Rank01Participants7Percentile100.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkQwenClawBenchScore61.8%Rank01Participants1Percentile100.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkQwenWorldBenchScore62.1%Rank01Participants2Percentile100.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkSimpleVQAScore0.817 pointsRank01Participants14Percentile100.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkSURDSScore77.2%Rank01Participants1Percentile100.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkVLADBenchScore77.2%Rank01Participants1Percentile100.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkWorldVQAScore61.1%Rank01Participants5Percentile100.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkApexScore22.7%Rank02Participants2Percentile0.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkClawEval-MMScore55.7%Rank02Participants4Percentile66.7%EvidenceCEvaluatedAug 17, 2026
BenchmarkIFEvalScore94.6%Rank02Participants67Percentile98.5%EvidenceCEvaluatedAug 17, 2026
BenchmarkMAXIFEScore88.8%Rank02Participants11Percentile90.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkMMLU-ProXScore85.4%Rank02Participants32Percentile96.8%EvidenceCEvaluatedAug 17, 2026
BenchmarkODinWScore51.1%Rank02Participants16Percentile93.3%EvidenceCEvaluatedAug 17, 2026
BenchmarkOmniDocBench 1.5Score91.4%Rank02Participants18Percentile94.1%EvidenceCEvaluatedAug 17, 2026
BenchmarkPolyMATHScore84.0%Rank02Participants23Percentile95.5%EvidenceCEvaluatedAug 17, 2026
BenchmarkRealWorldQAScore86.9%Rank02Participants29Percentile96.4%EvidenceCEvaluatedAug 17, 2026
BenchmarkTVBenchScore78.2%Rank02Participants3Percentile50.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkAndroidWorldScore81.0%Rank03Participants5Percentile50.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkBFCL-V4Score72.9%Rank03Participants15Percentile85.7%EvidenceCEvaluatedAug 17, 2026
BenchmarkMCP-MarkScore58.7%Rank03Participants8Percentile71.4%EvidenceCEvaluatedAug 17, 2026
BenchmarkNOVA-63Score58.8%Rank03Participants11Percentile80.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkSpreadSheetBench-v1Score86.3%Rank03Participants3Percentile0.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkSuperGPQAScore71.4%Rank03Participants34Percentile93.9%EvidenceCEvaluatedAug 17, 2026
BenchmarkVideo-MMEScore88.0%Rank03Participants17Percentile87.5%EvidenceCEvaluatedAug 17, 2026
BenchmarkVisFactorScore42.8%Rank03Participants3Percentile0.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkVITA-BenchScore45.6%Rank03Participants10Percentile77.8%EvidenceCEvaluatedAug 17, 2026
BenchmarkCoWorkBenchScore65.1%Rank04Participants4Percentile0.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkCritPTScore6.0%Rank04Participants5Percentile25.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkERQAScore69.8%Rank04Participants24Percentile87.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkGlobal PIQAScore90.3%Rank04Participants13Percentile75.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkHMMT Feb 26Score92.9%Rank04Participants12Percentile72.7%EvidenceCEvaluatedAug 17, 2026
BenchmarkMMLU-ProScore88.5%Rank04Participants134Percentile97.7%EvidenceCEvaluatedAug 17, 2026
BenchmarkMMLU-ReduxScore94.5%Rank04Participants48Percentile93.6%EvidenceCEvaluatedAug 17, 2026
BenchmarkSkillsBenchScore54.9%Rank04Participants8Percentile57.1%EvidenceCEvaluatedAug 17, 2026
BenchmarkWMT24++Score84.6%Rank04Participants23Percentile86.4%EvidenceCEvaluatedAug 17, 2026
BenchmarkBabyVisionScore70.4%Rank05Participants10Percentile55.6%EvidenceCEvaluatedAug 17, 2026
BenchmarkIncludeScore83.0%Rank05Participants31Percentile86.7%EvidenceCEvaluatedAug 17, 2026
BenchmarkLiveCodeBench v6Score89.6%Rank05Participants56Percentile92.7%EvidenceCEvaluatedAug 17, 2026
BenchmarkLVBenchScore76.2%Rank05Participants25Percentile83.3%EvidenceCEvaluatedAug 17, 2026
BenchmarkScreenSpot ProScore79.0%Rank05Participants25Percentile83.3%EvidenceCEvaluatedAug 17, 2026
BenchmarkVideoMMMUScore85.4%Rank05Participants26Percentile84.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkMathVisionScore90.3%Rank06Participants33Percentile84.4%EvidenceCEvaluatedAug 17, 2026
BenchmarkSciCodeScore51.3%Rank06Participants21Percentile75.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkIFBenchScore79.1%Rank08Participants34Percentile78.8%EvidenceCEvaluatedAug 17, 2026
BenchmarkIMO-AnswerBenchScore86.0%Rank08Participants20Percentile63.2%EvidenceCEvaluatedAug 17, 2026
BenchmarkTerminal-Bench 2.0Score70.3%Rank08Participants51Percentile86.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkClaw-EvalScore62.7%Rank09Participants14Percentile38.5%EvidenceCEvaluatedAug 17, 2026
BenchmarkSWE-bench MultilingualScore75.8%Rank11Participants38Percentile73.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkCharXiv-RScore85.9%Rank13Participants51Percentile76.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkNL2RepoScore41.1%Rank13Participants17Percentile25.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkMMMLUScore89.0%Rank14Participants49Percentile72.9%EvidenceCEvaluatedAug 17, 2026
BenchmarkOSWorld-VerifiedScore73.3%Rank14Participants24Percentile43.5%EvidenceCEvaluatedAug 17, 2026
BenchmarkFrontierCode 1.1Score10.2%Rank16Participants16Percentile0.0%EvidenceBEvaluatedAug 17, 2026
BenchmarkMMMU-ProScore79.0%Rank16Participants68Percentile77.6%EvidenceCEvaluatedAug 17, 2026
BenchmarkFinance Agent v2Score38.2%Rank18Participants26Percentile32.0%EvidenceBEvaluatedAug 17, 2026
BenchmarkMCP AtlasScore73.2%Rank18Participants33Percentile46.9%EvidenceCEvaluatedAug 17, 2026
BenchmarkGPQAScore90.3%Rank23Participants239Percentile90.8%EvidenceCEvaluatedAug 17, 2026
BenchmarkSWE-Bench ProScore57.6%Rank23Participants50Percentile55.1%EvidenceCEvaluatedAug 17, 2026
BenchmarkSWE-Bench VerifiedScore77.7%Rank25Participants111Percentile78.2%EvidenceCEvaluatedAug 17, 2026
BenchmarkHumanity's Last ExamScore34.7%Rank43Participants99Percentile57.1%EvidenceCEvaluatedAug 17, 2026

Qwen3.7 Plus Arena Results

Preference and agent-evaluation results for the default version.

13 rows
Columns

Show columns

Sort by
Arena
Category
Rank
Rating / score
Votes
Observations
Result date
Arenaagent bash recovery stepsCategoryoverallRank22Rating / score0.1VotesN/AObservations19.7KResult dateAug 13, 2026
ArenadocumentCategoryoverallRank25Rating / score1439.6Votes3,088ObservationsN/AResult dateJul 30, 2026
Arenadocument style controlCategoryoverallRank26Rating / score1448.5Votes3,088ObservationsN/AResult dateJul 30, 2026
ArenavisionCategoryoverallRank27Rating / score1277.7Votes7,200ObservationsN/AResult dateAug 6, 2026
ArenatextCategoryoverallRank31Rating / score1457.1Votes30,730ObservationsN/AResult dateAug 12, 2026
Arenavision style controlCategoryoverallRank31Rating / score1262.1Votes7,200ObservationsN/AResult dateAug 6, 2026
Arenaagent task outcome explicitCategoryoverallRank32Rating / score-0.0VotesN/AObservations11.3KResult dateAug 13, 2026
ArenaagentCategoryoverallRank34Rating / score-0.0VotesN/AObservations699.4KResult dateAug 13, 2026
Arenaagent steerabilityCategoryoverallRank38Rating / score-0.1VotesN/AObservations18.6KResult dateAug 13, 2026
Arenaagent tool hallucinationCategoryoverallRank38Rating / score0.0VotesN/AObservations645.6KResult dateAug 13, 2026
Arenaagent praise complaintCategoryoverallRank39Rating / score-0.1VotesN/AObservations4.3KResult dateAug 13, 2026
Arenatext factualityCategoryoverallRank45Rating / score1453.8Votes30,714ObservationsN/AResult dateAug 12, 2026
Arenatext style controlCategoryoverallRank48Rating / score1458.4Votes30,730ObservationsN/AResult dateAug 12, 2026

Qwen3.7 Plus Pricing

Official vendor API pricing appears first, followed by individual provider offers.

Official API
$0.50 input, $3 output per 1M
Official provider
Alibaba
Lowest third-party
From $0.282 input, $1.13 output per 1M via AIHubMix
Tracked offerings
27
27 rows
Columns

Show columns

Sort by
Provider
Provider model ID
Region
Input / 1M
Output / 1M
Context
Updated
ProviderAlibaba Token PlanProvider model IDqwen3.7-plusRegionglobalInput / 1MN/AOutput / 1MN/AContext1MUpdatedAug 17, 2026
ProviderKenariProvider model IDqwen3-7-plusRegionglobalInput / 1MN/AOutput / 1MN/AContext1MUpdatedAug 17, 2026
ProviderAlibaba Coding Plan (China)Provider model IDqwen3.7-plusRegionglobalInput / 1MN/AOutput / 1MN/AContext1MUpdatedAug 17, 2026
ProviderAlibaba Coding PlanProvider model IDqwen3.7-plusRegionglobalInput / 1MN/AOutput / 1MN/AContext1MUpdatedAug 17, 2026
ProviderAlibaba Token Plan (China)Provider model IDqwen3.7-plusRegionglobalInput / 1MN/AOutput / 1MN/AContext1MUpdatedAug 17, 2026
ProviderAIHubMixProvider model IDqwen3.7-plusRegionglobalInput / 1M$0.282Output / 1M$1.13Context991KUpdatedAug 17, 2026
ProviderCrossModelProvider model IDqwen/qwen3.7-plusRegionglobalInput / 1M$0.288Output / 1M$1.13Context1MUpdatedAug 17, 2026
ProviderOpenRouterProvider model IDqwen/qwen3.7-plusRegionglobalInput / 1M$0.32Output / 1M$1.28Context1MUpdatedAug 17, 2026
ProviderKilo GatewayProvider model IDqwen/qwen3.7-plusRegionglobalInput / 1M$0.32Output / 1M$1.28Context1MUpdatedAug 17, 2026
ProviderPioneerProvider model IDqwen3.7-plusRegionglobalInput / 1M$0.32Output / 1M$1.28Context1MUpdatedAug 17, 2026
ProviderNanoGPTProvider model IDqwen3.7-plusRegionglobalInput / 1M$0.40Output / 1M$1.6Context991.8KUpdatedAug 17, 2026
ProviderImpossiblProvider model IDqwen/qwen3.7-plusRegionglobalInput / 1M$0.40Output / 1M$1.6Context1MUpdatedAug 17, 2026
ProviderOpenCode GoProvider model IDqwen3.7-plusRegionglobalInput / 1M$0.40Output / 1M$1.6Context1MUpdatedAug 17, 2026
ProviderLLMTRProvider model IDqwen/qwen3.7-plusRegionglobalInput / 1M$0.40Output / 1M$1.6Context1MUpdatedAug 17, 2026
ProviderMerge GatewayProvider model IDqwen/qwen3.7-plusRegionglobalInput / 1M$0.40Output / 1M$1.6Context1MUpdatedAug 17, 2026
ProviderZenMuxProvider model IDqwen/qwen3.7-plusRegionglobalInput / 1M$0.40Output / 1M$1.6Context1MUpdatedAug 17, 2026
ProviderOfoxProvider model IDbailian/qwen3.7-plusRegionglobalInput / 1M$0.40Output / 1M$1.6Context1MUpdatedAug 17, 2026
ProviderClinePassProvider model IDcline-pass/qwen3.7-plusRegionglobalInput / 1M$0.40Output / 1M$1.6Context1MUpdatedAug 17, 2026
ProviderEmpirioLabs AIProvider model IDqwen3-7-plusRegionglobalInput / 1M$0.40Output / 1M$1.6Context1MUpdatedAug 17, 2026
ProviderVercel AI GatewayProvider model IDalibaba/qwen3.7-plusRegionglobalInput / 1M$0.40Output / 1M$1.6Context1MUpdatedAug 17, 2026
ProviderLLM GatewayProvider model IDqwen3.7-plusRegionglobalInput / 1M$0.40Output / 1M$1.6Context1MUpdatedAug 17, 2026
ProviderFireworks AIProvider model IDaccounts/fireworks/models/qwen3p7-plusRegionglobalInput / 1M$0.40Output / 1M$1.6Context262.1KUpdatedAug 17, 2026
ProviderAlibaba (China)Provider model IDqwen3.7-plusRegionglobalInput / 1M$0.50Output / 1M$3Context1MUpdatedAug 17, 2026
ProviderAlibabaProvider model IDqwen3.7-plusRegionglobalInput / 1M$0.50Output / 1M$3Context1MUpdatedAug 17, 2026
ProviderVenice AIProvider model IDqwen-3-7-plusRegionglobalInput / 1M$0.50Output / 1M$2Context1MUpdatedAug 17, 2026
ProviderModelisProvider model IDqwen/qwen3.7-plusRegionglobalInput / 1M$0.768Output / 1M$3.07Context1MUpdatedAug 17, 2026
ProviderCharm HyperProvider model IDqwen3.7-plusRegionglobalInput / 1M$1.2Output / 1M$4.8Context1MUpdatedAug 17, 2026

Official prices use only the vendor's configured official Provider and positive standard USD PAYG rates. Third-party offers remain explicitly labeled.

Qwen3.7 Plus Runtime Performance

Provider-specific output speed and catalog latency for Qwen3.7-Plus. Runtime does not affect the capability score.

No runtime data

No provider-specific speed or latency record is linked to the default version yet.

Browse runtime rankings

Output Speed is generated output tokens received per second. Catalog latency is reported separately from observed provider TTFT.

Qwen3.7 Plus Specifications

Technical details for the model's default version.

Version
Qwen3.7-Plus
Released
May 31, 2026
Knowledge cutoff
Apr 1, 2025
Parameters
N/A
Context window
1M
Max output
64K
Inputs
image, text, video
Outputs
text
Open weights
No
License
Proprietary

Qwen3.7 Plus Versions

Available versions of this model. The score column identifies the version used in the overall ranking.

1 rows
Columns

Show columns

Sort by
Version
Released
LLMBoard
Parameters
Context
Max output
Open weights
License
VersionQwen3.7-PlusReleasedMay 31, 2026LLMBoard73.2ParametersN/AContext1MMax output64KOpen weightsNoLicenseProprietary

Qwen3.7 Plus vs nearby models

Open a comparison with the three ranked models immediately above and below this model.

Qwen3.7 PlusvsGemini 3.6 FlashQwen3.7 PlusvsSeed 2.1 TurboQwen3.7 PlusvsMuse SparkQwen3.7 PlusvsClaude Sonnet 4.6Qwen3.7 PlusvsSakana NamazuQwen3.7 PlusvsGPT-5.2-Pro

Models similar to Qwen3.7 Plus

Recommendations prioritize the same model type and family, then the closest LLMBoard score.

#24+4.5
AC

Qwen3.8 27B

Alibaba Cloud / Qwen Team

77.6 LLMBoard

DetailsCompare
#23+5.1
AC

Qwen3.7 Max

Alibaba Cloud / Qwen Team

78.3 LLMBoard

DetailsCompare
#48-6.1
AC

Qwen3.6 Plus

Alibaba Cloud / Qwen Team

67.0 LLMBoard

DetailsCompare
#51-7.3
AC

Qwen3.5 397B A17B

Alibaba Cloud / Qwen Team

65.9 LLMBoard

DetailsCompare
#65-13.2
AC

Qwen3.6 27B

Alibaba Cloud / Qwen Team

59.9 LLMBoard

DetailsCompare
#70-14.3
AC

Qwen3.5 122B A10B

Alibaba Cloud / Qwen Team

58.9 LLMBoard

DetailsCompare

What is Qwen3.7 Plus?

Key information about Qwen3.7 Plus and its available data.

7-Plus is Alibaba Cloud Qwen Team's multimodal agent model that unifies vision and language into a single agent foundation.

7 text backbone, it operates as a multimodal interactive hybrid agent—perceiving real-world scenes, reading screens and operating GUIs, writing code from visual references, navigating mobile apps end-to-end, and answering search-augmented visual questions—while blending GUI and CLI interactions within a single agent loop.

It is a versatile coding agent and productivity assistant with full-modality input, generalizing across scaffolds such as Claude Code, OpenClaw, and Qwen Code. Features a 1 million token context window, up to 65,536 output tokens, always-on thinking, and a preserve_thinking mode for agentic tasks.

Available via Alibaba Cloud Model Studio (DashScope).

Data as of 2026-08-17.

FAQ

Common questions about Qwen3.7 Plus.

When was Qwen3.7 Plus released?

Qwen3.7 Plus's default version was released on May 31, 2026.

How much does Qwen3.7 Plus cost?

Qwen3.7 Plus's official API price is $0.50 per million input tokens and $3 per million output tokens via Alibaba. The lowest tracked third-party offer starts at $0.282 input and $1.13 output via AIHubMix.

Who created Qwen3.7 Plus?

Qwen3.7 Plus was created by Alibaba Cloud / Qwen Team.

What is the context window for Qwen3.7 Plus?

The default version has a 1M token context window.

Is Qwen3.7 Plus open weight?

No. The default version is not marked as having publicly available weights.

How many API providers offer Qwen3.7 Plus?

27 provider offerings are linked to the default version.

What models should I compare Qwen3.7 Plus with?

Nearby ranked alternatives include Gemini 3.6 Flash, Seed 2.1 Turbo, Muse Spark.