llmboard.aiLeaderboard Center
Overall
Overall RankingOpen Models
Tools
Model DirectoryCompare Models
Capabilities
CodingReasoningMathKnowledgeInstruction Following
Price & Efficiency
Price & ValueCapability vs. PriceRuntime Performance
Modalities
Image GenerationVideo GenerationSpeech ModelsEmbeddings
Core Benchmarks
GPQAMMLU-ProAIME 2025SWE-Bench VerifiedMMLUHumanity's Last ExamLiveCodeBenchMATHHumanEvalMMMU-ProView all benchmarks
Methods
Scoring & Data
393 models668 benchmarks

Leaderboard Center

Overall RankingCodingCore BenchmarksPrice & ValueRuntime Performance

Modalities

All ModelsImage GenerationImage EditingVideo GenerationImage-to-VideoVideo EditingText-to-SpeechSpeech-to-TextEmbeddings

Data & Methods

Scoring MethodAll BenchmarksReasoningMath

Vendors

All VendorsOpenAIAnthropicGoogle
llmboard.aiCopyright 2026 llmboard.ai

Inception model product

Mercury 2

Mercury 2 is the fastest reasoning LLM, built on diffusion-based language model (dLLM) architecture.

Updated Aug 12, 2026. Default version: Mercury 2

Compare
LLMBoard score39.3Mercury 2
Coverage80%6 benchmark families
Context window128KTokens
Official input price$0.25Inception API

On this page

  • Specification
  • Capability
  • Benchmarks
  • Arena
  • Runtime
  • Pricing
  • Versions
  • About
  • Compare
  • Similar models
  • FAQ

Mercury 2 Specifications

Technical details for the model's default version.

Version
Mercury 2
Released
Feb 24, 2026
Knowledge cutoff
Unknown
Parameters
N/A
Context window
128K
Max output
8.2K
Inputs
text
Outputs
text
Open weights
No
License
Proprietary

Mercury 2 Capability Profile

This profile uses the model's current scored version. Arena ratings and prices are shown separately.

Mercury 2 category scores

Mercury 2 Benchmark Results

Benchmark scores for Mercury 2.

6 rows
Columns

Show columns

IFBench71.0%152950.0%CAug 11, 2026
SciCode38.0%161916.7%CAug 11, 2026
Tau2 Airline53.0%192318.2%CAug 11, 2026
LiveCodeBench67.0%247368.1%CAug 11, 2026
AIME 202591.1%4311462.8%CAug 11, 2026
GPQA74.0%11423451.5%CAug 11, 2026

Mercury 2 Arena Results

Preference and agent-evaluation results for the default version.

4 rows
Columns

Show columns

webdevoverall1081166.2908N/AAug 10, 2026
text factualityoverall1451371.93,126N/AAug 10, 2026
textoverall1731357.33,127N/AAug 10, 2026
text style controloverall1981346.43,127N/AAug 10, 2026

Mercury 2 Runtime Performance

Provider-specific output speed and catalog latency for Mercury 2. Runtime does not affect the capability score.

1 rows
Columns

Show columns

Inception437.018 tok/s2.958 s128K8.2KAug 11, 2026

Output Speed is generated output tokens received per second. Catalog latency is reported separately from observed provider TTFT.

Mercury 2 Pricing

Official vendor API pricing appears first, followed by individual provider offers.

Official API
$0.25 input, $0.75 output per 1M
Official provider
Inception
Lowest third-party
From $0.25 input, $0.75 output per 1M via NanoGPT
Tracked offerings
5
5 rows
Columns

Show columns

NanoGPTmercury-2global$0.25$0.75128KAug 11, 2026
Inceptionmercury-2global$0.25$0.75128KAug 11, 2026
Vercel AI Gatewayinception/mercury-2global$0.25$0.75128KAug 11, 2026
OpenRouterinception/mercury-2global$0.25$0.75128KAug 11, 2026
Venice AImercury-2global$0.3125$0.9375128KAug 11, 2026

Official prices use only the vendor's configured official Provider and positive standard USD PAYG rates. Third-party offers remain explicitly labeled.

Mercury 2 Versions

Available versions of this model. The score column identifies the version used in the overall ranking.

1 rows
Columns

Show columns

Mercury 2Feb 24, 202639.3N/A128K8.2KNoProprietary

What is Mercury 2?

Key information about Mercury 2 and its available data.

Mercury 2 is the fastest reasoning LLM, built on diffusion-based language model (dLLM) architecture. Instead of generating text token-by-token, it refines multiple text blocks simultaneously, achieving over 1,000 tokens per second on Nvidia Blackwell GPUs — 5x faster than leading speed-optimized LLMs.

Supports tool usage and JSON output with 128K context window.

Data as of 2026-08-11.

Mercury 2 vs nearby models

Open a comparison with the three ranked models immediately above and below this model.

Mercury 2vsQwen3 VL 32B ThinkingMercury 2vsDeepSeek-V3.1Mercury 2vsClaude Sonnet 4Mercury 2vsClaude Sonnet 3.7Mercury 2vsGLM 4.7 FlashMercury 2vsGemma 4 12B

Models similar to Mercury 2

Recommendations prioritize the same model type and family, then the closest LLMBoard score.

#129+0.2
AN

Claude Sonnet 4

Anthropic

39.5 LLMBoard

DetailsCompare
#128+0.6
DE

DeepSeek-V3.1

DeepSeek

39.9 LLMBoard

DetailsCompare
#131-0.6
AN

Claude Sonnet 3.7

Anthropic

38.7 LLMBoard

DetailsCompare
#132-0.9
ZA

GLM 4.7 Flash

Zhipu AI

38.5 LLMBoard

DetailsCompare
#133-1.1
GO

Gemma 4 12B

Google

38.2 LLMBoard

DetailsCompare
#127+1.2
AC

Qwen3 VL 32B Thinking

Alibaba Cloud / Qwen Team

40.5 LLMBoard

DetailsCompare

FAQ

Common questions about Mercury 2.

When was Mercury 2 released?

Mercury 2's default version was released on Feb 24, 2026.

How much does Mercury 2 cost?

Mercury 2's official API price is $0.25 per million input tokens and $0.75 per million output tokens via Inception. The lowest tracked third-party offer starts at $0.25 input and $0.75 output via NanoGPT.

Who created Mercury 2?

Mercury 2 was created by Inception.

What is the context window for Mercury 2?

The default version has a 128K token context window.

Is Mercury 2 open weight?

No. The default version is not marked as having publicly available weights.

How many API providers offer Mercury 2?

5 provider offerings are linked to the default version.

What models should I compare Mercury 2 with?

Nearby ranked alternatives include Qwen3 VL 32B Thinking, DeepSeek-V3.1, Claude Sonnet 4.