llmboard.aiLeaderboard Center
Overall
Overall RankingOpen Models
Tools
Model DirectoryCompare Models
Capabilities
CodingReasoningMathKnowledgeInstruction Following
Price & Efficiency
Price & ValueCapability vs. PriceRuntime Performance
Modalities
Image GenerationVideo GenerationSpeech ModelsEmbeddings
Core Benchmarks
GPQAMMLU-ProAIME 2025SWE-Bench VerifiedMMLUHumanity's Last ExamLiveCodeBenchMATHHumanEvalMMMU-ProView all benchmarks
Methods
Scoring & Data
393 models668 benchmarks

Leaderboard Center

Overall RankingCodingCore BenchmarksPrice & ValueRuntime Performance

Modalities

All ModelsImage GenerationImage EditingVideo GenerationImage-to-VideoVideo EditingText-to-SpeechSpeech-to-TextEmbeddings

Data & Methods

Scoring MethodAll BenchmarksReasoningMath

Vendors

All VendorsOpenAIAnthropicGoogle
llmboard.aiCopyright 2026 llmboard.ai

Tencent model product

Hy3

8B MTP layer, developed by the Tencent Hy Team.

Updated Aug 12, 2026. Default version: Hy3

Compare
LLMBoard score71.0Hy3
Coverage80%24 benchmark families
Context window256KTokens
Official input priceN/AOfficial price unavailable

On this page

  • Specification
  • Capability
  • Benchmarks
  • Arena
  • Runtime
  • Pricing
  • Versions
  • About
  • Compare
  • Similar models
  • FAQ

Hy3 Specifications

Technical details for the model's default version.

Version
Hy3
Released
Jul 6, 2026
Knowledge cutoff
Unknown
Parameters
295B
Context window
256K
Max output
64K
Inputs
text
Outputs
text
Open weights
Yes
License
Apache 2.0

Hy3 Capability Profile

This profile uses the model's current scored version. Arena ratings and prices are shown separately.

Hy3 category scores

Hy3 Benchmark Results

Benchmark scores for Hy3.

31 rows
Columns

Show columns

CL-bench23.8%012100.0%CAug 11, 2026
CL-bench (Life)17.0%011100.0%CAug 11, 2026
CMT-Benchmark37.9%011100.0%CAug 11, 2026
HorizonMath7.1%013100.0%CAug 11, 2026
Humanity's Last Exam (no tools, text-only)47.0%011100.0%CAug 11, 2026
PHYBench77.4%011100.0%CAug 11, 2026
AA-LCR73.4%021693.3%CAug 11, 2026
ArXivMath52.2%0220.0%CAug 11, 2026
Humanity's Last Exam (with tools, text-only)53.2%0220.0%CAug 11, 2026
FrontierScience Olympiad74.8%0330.0%CAug 11, 2026
IMO-AnswerBench90.0%031988.9%CAug 11, 2026
SkillsBench55.3%03871.4%CAug 11, 2026
SuperChem54.9%0330.0%CAug 11, 2026
USAMO 202630.24 points0330.0%CAug 11, 2026
WideSearch76.4%03975.0%CAug 11, 2026
WildClawBench53.6%03550.0%CAug 11, 2026
Claw-Eval68.5%041375.0%CAug 11, 2026
DeepSearchQA91.0%04962.5%CAug 11, 2026
FrontierScience Research21.3%0440.0%CAug 11, 2026
MathArena Apex38.7%04750.0%CAug 11, 2026
NL2Repo45.6%061461.5%CAug 11, 2026
APEX-Agents25.6%0770.0%CAug 11, 2026
MCP Atlas79.1%073180.0%CAug 11, 2026
DeepSWE28.0%091011.1%CAug 11, 2026
SWE-bench Multilingual75.8%093475.8%CAug 11, 2026
Terminal-Bench 2.171.7%131933.3%CAug 11, 2026
BrowseComp84.2%145877.2%CAug 11, 2026
Toolathlon48.5%173146.7%CAug 11, 2026
SWE-Bench Pro57.9%194559.1%CAug 11, 2026
SWE-Bench Verified78.0%2010581.7%CAug 11, 2026
GPQA90.4%2123491.4%CAug 11, 2026

Hy3 Arena Results

Preference and agent-evaluation results for the default version.

14 rows
Columns

Show columns

agent praise complaintoverall190.0N/A2KAug 6, 2026
webdevoverall201524.22,175N/AAug 10, 2026
agent bash recovery stepsoverall280.0N/A12.6KAug 6, 2026
agentoverall30-0.0N/A477.6KAug 6, 2026
agent task outcome explicitoverall31-0.0N/A6.7KAug 6, 2026
agent steerabilityoverall39-0.1N/A8.1KAug 6, 2026
agent tool hallucinationoverall41-0.0N/A448.2KAug 6, 2026
text style controloverall521456.34,544N/AAug 10, 2026
textoverall581441.04,544N/AAug 10, 2026
text factualityoverall591448.04,544N/AAug 10, 2026
webdevoverall791356.31,394N/AAug 10, 2026
text factualityoverall1171415.76,624N/AAug 10, 2026
text style controloverall1221412.86,627N/AAug 10, 2026
textoverall1271405.06,627N/AAug 10, 2026

Hy3 Runtime Performance

Provider-specific output speed and catalog latency for Hy3. Runtime does not affect the capability score.

No runtime data

No provider-specific speed or latency record is linked to the default version yet.

Browse runtime rankings

Output Speed is generated output tokens received per second. Catalog latency is reported separately from observed provider TTFT.

Hy3 Pricing

Official vendor API pricing appears first, followed by individual provider offers.

Official API
N/A
Official provider
N/A
Lowest third-party
From $0.066 input, $0.26 output per 1M via NanoGPT
Tracked offerings
11
11 rows
Columns

Show columns

Tencent TokenHubhy3globalN/AN/A256KAug 11, 2026
Tencent Token Planhy3globalN/AN/A256KAug 11, 2026
NanoGPTtencent/hy3global$0.066$0.26262.1KAug 11, 2026
OpenRoutertencent/hy3global$0.132$0.528262.1KAug 11, 2026
OpenCode Gohy3global$0.14$0.58256KAug 11, 2026
LLM Gatewayhy3global$0.14$0.58262.1KAug 11, 2026
Kilo Gatewaytencent/hy3global$0.14$0.58262.1KAug 11, 2026
Vercel AI Gatewaytencent/hy3global$0.14$0.58262.1KAug 11, 2026
Hugging Facetencent/Hy3global$0.14$0.58262.1KAug 11, 2026
Deep Infratencent/Hy3global$0.14$0.58262.1KAug 11, 2026
CrossModeltencent/hy3global$0.16$0.64262.1KAug 11, 2026

Official prices use only the vendor's configured official Provider and positive standard USD PAYG rates. Third-party offers remain explicitly labeled.

Hy3 Versions

Available versions of this model. The score column identifies the version used in the overall ranking.

1 rows
Columns

Show columns

Hy3Jul 6, 202671.0295B256K64KYesApache 2.0

What is Hy3?

Key information about Hy3 and its available data.

8B MTP layer, developed by the Tencent Hy Team. Following the Hy3 Preview launch in late April, the team scaled up post-training with higher-quality data and RL, gathering feedback from 50+ products.

Hy3 outperforms similar-size models and rivals flagship open-source models with 2-5x the parameters, with strong gains in reasoning, agentic, and long-context tasks.

It uses 80 layers (plus 1 MTP layer), 64 GQA attention heads (8 KV heads, head dim 128), a 4096 hidden size, 192 experts with top-8 activated, a 256K context window, and BF16 precision.

Hy3 is a hybrid-thinking model supporting configurable reasoning effort (no_think, low, high), and emphasizes production-grade tool-call and output-format stability, reduced hallucination, and reliable multi-turn intent tracking.

Data as of 2026-08-11.

Hy3 vs nearby models

Open a comparison with the three ranked models immediately above and below this model.

Hy3vsGPT-5.2-ProHy3vsGPT-5.2Hy3vsMiniMax M3Hy3vsGLM 5.1Hy3vsSeed 2.0 ProHy3vsClaude Opus 4.5

Models similar to Hy3

Recommendations prioritize the same model type and family, then the closest LLMBoard score.

#34+0.3
MI

MiniMax M3

MiniMax

71.3 LLMBoard

DetailsCompare
#33+0.4
OP

GPT-5.2

OpenAI

71.4 LLMBoard

DetailsCompare
#32+0.6
OP

GPT-5.2-Pro

OpenAI

71.6 LLMBoard

DetailsCompare
#36-1.4
ZA

GLM 5.1

Zhipu AI

69.6 LLMBoard

DetailsCompare
#37-1.5
BY

Seed 2.0 Pro

ByteDance

69.5 LLMBoard

DetailsCompare
#38-1.5
AN

Claude Opus 4.5

Anthropic

69.5 LLMBoard

DetailsCompare

FAQ

Common questions about Hy3.

When was Hy3 released?

Hy3's default version was released on Jul 6, 2026.

How much does Hy3 cost?

No official standard PAYG price is currently available for Hy3. The lowest tracked third-party offer starts at $0.066 input and $0.26 output via NanoGPT.

Who created Hy3?

Hy3 was created by Tencent.

What is the context window for Hy3?

The default version has a 256K token context window.

Is Hy3 open weight?

Yes. The default version is marked as open weight under Apache 2.0.

How many API providers offer Hy3?

11 provider offerings are linked to the default version.

What models should I compare Hy3 with?

Nearby ranked alternatives include GPT-5.2-Pro, GPT-5.2, MiniMax M3.