Alibaba Cloud / Qwen Team model product
7-Plus is Alibaba Cloud Qwen Team's multimodal agent model that unifies vision and language into a single agent foundation.
Updated Aug 12, 2026. Default version: Qwen3.7-Plus
Technical details for the model's default version.
This profile uses the model's current scored version. Arena ratings and prices are shown separately.
Benchmark scores for Qwen3.7-Plus.
| BC-VL | 51.1% | 01 | 1 | 100.0% | C | |
| ClawEval-MM | 55.7% | 01 | 3 | 100.0% | C | |
| CountQA | 77.0% | 01 | 1 | 100.0% | C | |
| DeepPlanning | 62.3% | 01 | 9 | 100.0% | C | |
| HiPhO | 84.1% | 01 | 1 | 100.0% | C | |
| LingoQA | 83.4% | 01 | 4 | 100.0% | C | |
| MedXpertQA-MM | 71.0% | 01 | 1 | 100.0% | C | |
| MLVU | 87.4% | 01 | 10 | 100.0% | C | |
| MMBC | 46.3% | 01 | 1 | 100.0% | C | |
| MMSearch-Plus | 41.4% | 01 | 2 | 100.0% | C | |
| MRCR v2 | 91.7% | 01 | 3 | 100.0% | C | |
| OCRBench_V2 | 67.1% | 01 | 7 | 100.0% | C | |
| QwenClawBench | 61.8% | 01 | 1 | 100.0% | C | |
| QwenWorldBench | 62.1% | 01 | 2 | 100.0% | C | |
| SimpleVQA | 0.817 points | 01 | 13 | 100.0% | C | |
| SURDS | 77.2% | 01 | 1 | 100.0% | C | |
| VLADBench | 77.2% | 01 | 1 | 100.0% | C | |
| WorldVQA | 61.1% | 01 | 5 | 100.0% | C | |
| AndroidWorld | 81.0% | 02 | 4 | 66.7% | C | |
| Apex | 22.7% | 02 | 2 | 0.0% | C | |
| IFEval | 94.6% | 02 | 65 | 98.4% | C | |
| MAXIFE | 88.8% | 02 | 11 | 90.0% | C | |
| MMLU-ProX | 85.4% | 02 | 32 | 96.8% | C | |
| ODinW | 51.1% | 02 | 16 | 93.3% | C | |
| OmniDocBench 1.5 | 91.4% | 02 | 17 | 93.8% | C | |
| PolyMATH | 84.0% | 02 | 23 | 95.5% | C | |
| RealWorldQA | 86.9% | 02 | 26 | 96.0% | C | |
| TVBench | 78.2% | 02 | 3 | 50.0% | C | |
| BFCL-V4 | 72.9% | 03 | 13 | 83.3% | C | |
| CoWorkBench | 65.1% | 03 | 3 | 0.0% | C | |
| CritPT | 6.0% | 03 | 4 | 33.3% | C | |
| LiveCodeBench v6 | 89.6% | 03 | 53 | 96.2% | C | |
| MCP-Mark | 58.7% | 03 | 8 | 71.4% | C | |
| MMLU-Pro | 88.5% | 03 | 129 | 98.4% | C | |
| NOVA-63 | 58.8% | 03 | 11 | 80.0% | C | |
| SpreadSheetBench-v1 | 86.3% | 03 | 3 | 0.0% | C | |
| SuperGPQA | 71.4% | 03 | 34 | 93.9% | C | |
| Video-MME | 88.0% | 03 | 17 | 87.5% | C | |
| VisFactor | 42.8% | 03 | 3 | 0.0% | C | |
| VITA-Bench | 45.6% | 03 | 10 | 77.8% | C | |
| BabyVision | 70.4% | 04 | 9 | 62.5% | C | |
| ERQA | 69.8% | 04 | 23 | 86.4% | C | |
| Global PIQA | 90.3% | 04 | 13 | 75.0% | C | |
| HMMT Feb 26 | 92.9% | 04 | 11 | 70.0% | C | |
| LVBench | 76.2% | 04 | 24 | 87.0% | C | |
| MMLU-Redux | 94.5% | 04 | 48 | 93.6% | C | |
| SkillsBench | 54.9% | 04 | 8 | 57.1% | C | |
| WMT24++ | 84.6% | 04 | 23 | 86.4% | C | |
| Include | 83.0% | 05 | 31 | 86.7% | C | |
| MathVision | 90.3% | 05 | 32 | 87.1% | C | |
| ScreenSpot Pro | 79.0% | 05 | 25 | 83.3% | C | |
| VideoMMMU | 85.4% | 05 | 26 | 84.0% | C | |
| IFBench | 79.1% | 06 | 29 | 82.1% | C | |
| SciCode | 51.3% | 06 | 19 | 72.2% | C | |
| Claw-Eval | 62.7% | 08 | 13 | 41.7% | C | |
| IMO-AnswerBench | 86.0% | 08 | 19 | 61.1% | C | |
| Terminal-Bench 2.0 | 70.3% | 08 | 49 | 85.4% | C | |
| NL2Repo | 41.1% | 10 | 14 | 30.8% | C | |
| SWE-bench Multilingual | 75.8% | 10 | 34 | 72.7% | C | |
| CharXiv-R | 85.9% | 11 | 48 | 78.7% | C | |
| OSWorld-Verified | 73.3% | 13 | 23 | 45.5% | C | |
| MMMLU | 89.0% | 14 | 49 | 72.9% | C | |
| FrontierCode 1.1 | 10.2% | 15 | 15 | 0.0% | B | |
| MMMU-Pro | 79.0% | 16 | 66 | 76.9% | C | |
| MCP Atlas | 73.2% | 17 | 31 | 46.7% | C | |
| Finance Agent v2 | 38.2% | 18 | 26 | 32.0% | B | |
| SWE-Bench Pro | 57.6% | 21 | 45 | 54.5% | C | |
| GPQA | 90.3% | 23 | 234 | 90.6% | C | |
| SWE-Bench Verified | 77.7% | 23 | 105 | 78.8% | C | |
| Humanity's Last Exam | 34.7% | 40 | 93 | 57.6% | C |
Preference and agent-evaluation results for the default version.
| agent bash recovery steps | overall | 24 | 0.1 | N/A | 18.4K | |
| document | overall | 25 | 1439.6 | 3,088 | N/A | |
| document style control | overall | 26 | 1448.5 | 3,088 | N/A | |
| agent task outcome explicit | overall | 27 | -0.0 | N/A | 10.7K | |
| vision | overall | 27 | 1277.7 | 7,200 | N/A | |
| agent | overall | 31 | -0.0 | N/A | 653.6K | |
| text | overall | 31 | 1456.4 | 30,246 | N/A | |
| vision style control | overall | 31 | 1262.1 | 7,200 | N/A | |
| agent steerability | overall | 32 | -0.0 | N/A | 17.4K | |
| agent tool hallucination | overall | 36 | 0.0 | N/A | 603.1K | |
| agent praise complaint | overall | 38 | -0.1 | N/A | 4K | |
| text factuality | overall | 45 | 1453.0 | 30,230 | N/A | |
| text style control | overall | 50 | 1457.6 | 30,246 | N/A |
Provider-specific output speed and catalog latency for Qwen3.7-Plus. Runtime does not affect the capability score.
| Together | 84.543 tok/s | 2.994 s | 1M | 65.5K | |
| Fireworks | 28.465 tok/s | N/A | 262.1K | 65.5K |
Output Speed is generated output tokens received per second. Catalog latency is reported separately from observed provider TTFT.
Official vendor API pricing appears first, followed by individual provider offers.
| Alibaba Coding Plan (China) | qwen3.7-plus | global | N/A | N/A | 1M | |
| Kenari | qwen3-7-plus | global | N/A | N/A | 1M | |
| Alibaba Coding Plan | qwen3.7-plus | global | N/A | N/A | 1M | |
| Alibaba Token Plan | qwen3.7-plus | global | N/A | N/A | 1M | |
| Alibaba Token Plan (China) | qwen3.7-plus | global | N/A | N/A | 1M | |
| AIHubMix | qwen3.7-plus | global | $0.282 | $1.13 | 991K | |
| CrossModel | qwen/qwen3.7-plus | global | $0.288 | $1.13 | 1M | |
| Pioneer | qwen3.7-plus | global | $0.32 | $1.28 | 1M | |
| Kilo Gateway | qwen/qwen3.7-plus | global | $0.32 | $1.28 | 1M | |
| OpenRouter | qwen/qwen3.7-plus | global | $0.32 | $1.28 | 1M | |
| Impossibl | qwen/qwen3.7-plus | global | $0.40 | $1.6 | 1M | |
| Fireworks AI | accounts/fireworks/models/qwen3p7-plus | global | $0.40 | $1.6 | 262.1K | |
| OpenCode Go | qwen3.7-plus | global | $0.40 | $1.6 | 1M | |
| NanoGPT | qwen3.7-plus | global | $0.40 | $1.6 | 991.8K | |
| LLM Gateway | qwen3.7-plus | global | $0.40 | $1.6 | 1M | |
| ClinePass | cline-pass/qwen3.7-plus | global | $0.40 | $1.6 | 1M | |
| ZenMux | qwen/qwen3.7-plus | global | $0.40 | $1.6 | 1M | |
| EmpirioLabs AI | qwen3-7-plus | global | $0.40 | $1.6 | 1M | |
| Ofox | bailian/qwen3.7-plus | global | $0.40 | $1.6 | 1M | |
| Merge Gateway | qwen/qwen3.7-plus | global | $0.40 | $1.6 | 1M | |
| Vercel AI Gateway | alibaba/qwen3.7-plus | global | $0.40 | $1.6 | 1M | |
| Venice AI | qwen-3-7-plus | global | $0.50 | $2 | 1M | |
| Alibaba | qwen3.7-plus | global | $0.50 | $3 | 1M | |
| Alibaba (China) | qwen3.7-plus | global | $0.50 | $3 | 1M | |
| Modelis | qwen/qwen3.7-plus | global | $0.768 | $3.07 | 1M | |
| Charm Hyper | qwen3.7-plus | global | $1.2 | $4.8 | 1M |
Official prices use only the vendor's configured official Provider and positive standard USD PAYG rates. Third-party offers remain explicitly labeled.
Available versions of this model. The score column identifies the version used in the overall ranking.
| Qwen3.7-Plus | 73.8 | N/A | 1M | 64K | No | Proprietary |
Key information about Qwen3.7 Plus and its available data.
7-Plus is Alibaba Cloud Qwen Team's multimodal agent model that unifies vision and language into a single agent foundation.
7 text backbone, it operates as a multimodal interactive hybrid agent—perceiving real-world scenes, reading screens and operating GUIs, writing code from visual references, navigating mobile apps end-to-end, and answering search-augmented visual questions—while blending GUI and CLI interactions within a single agent loop.
It is a versatile coding agent and productivity assistant with full-modality input, generalizing across scaffolds such as Claude Code, OpenClaw, and Qwen Code. Features a 1 million token context window, up to 65,536 output tokens, always-on thinking, and a preserve_thinking mode for agentic tasks.
Available via Alibaba Cloud Model Studio (DashScope).
Data as of 2026-08-11.
Open a comparison with the three ranked models immediately above and below this model.
Recommendations prioritize the same model type and family, then the closest LLMBoard score.
Common questions about Qwen3.7 Plus.
Qwen3.7 Plus's default version was released on May 31, 2026.
Qwen3.7 Plus's official API price is $0.50 per million input tokens and $3 per million output tokens via Alibaba. The lowest tracked third-party offer starts at $0.282 input and $1.13 output via AIHubMix.
Qwen3.7 Plus was created by Alibaba Cloud / Qwen Team.
The default version has a 1M token context window.
No. The default version is not marked as having publicly available weights.
26 provider offerings are linked to the default version.
Nearby ranked alternatives include GPT-5.5-Pro, Kimi K2.6, DeepSeek-V4-Pro.