llmboard.aiAI model intelligence
Home

Model Rankings

OverallOpen ModelsAgentCodingReasoningMathKnowledgeInstruction FollowingTextVision
Image GenerationImage Editing
Video GenerationImage to VideoVideo Editing
Text to SpeechSpeech to Text
Embeddings

Efficiency

Chat Token PricingImage PricingVideo PricingAudio Pricing
Chat Speed & LatencyProvider Reliability

Benchmarks

GPQAMMLU-ProAIME 2025SWE-Bench VerifiedMMLUHumanity's Last ExamLiveCodeBenchMATHHumanEvalMMMU-Pro
All Benchmarks

Tools

Model DirectoryCompare Models

Scoring & Data

Scoring & Data
393 models668 benchmarks

Leaderboard Center

Overall RankingCodingCore BenchmarksPrice & ValueRuntime Performance

Modalities

All ModelsImage GenerationImage EditingVideo GenerationImage-to-VideoVideo EditingText-to-SpeechSpeech-to-TextEmbeddings

Data & Methods

Scoring MethodAll BenchmarksReasoningMath

Vendors

All VendorsOpenAIAnthropicGoogle
llmboard.aiCopyright 2026 llmboard.ai

Tencent model product

Hy3

8B MTP layer, developed by the Tencent Hy Team.

Updated Aug 17, 2026. Default version: Hy3

Compare
LLMBoard Score70.7Hy3
Coverage80%25 benchmark families
Context window256KTokens
Official input priceN/AOfficial price unavailable

On this page

  • Capability
  • Benchmarks
  • Arena
  • Pricing
  • Runtime
  • Specification
  • Versions
  • Compare
  • Similar models
  • About
  • FAQ

Hy3 Capability Profile

This profile uses the model's current scored version. Arena ratings and prices are shown separately.

Hy3 LLMBoard score breakdown

Hy3 Benchmark Results

Benchmark scores for Hy3.

30 of 31 rows
Columns

Show columns

Sort by
Benchmark
Score
Rank
Participants
Percentile
Evidence
Evaluated
BenchmarkCL-benchScore23.8%Rank01Participants2Percentile100.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkCL-bench (Life)Score17.0%Rank01Participants1Percentile100.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkCMT-BenchmarkScore37.9%Rank01Participants1Percentile100.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkHorizonMathScore7.1%Rank01Participants3Percentile100.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkHumanity's Last Exam (no tools, text-only)Score47.0%Rank01Participants2Percentile100.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkPHYBenchScore77.4%Rank01Participants1Percentile100.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkAA-LCRScore73.4%Rank02Participants18Percentile94.1%EvidenceCEvaluatedAug 17, 2026
BenchmarkArXivMathScore52.2%Rank02Participants2Percentile0.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkFrontierScience OlympiadScore74.8%Rank03Participants3Percentile0.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkHumanity's Last Exam (with tools, text-only)Score53.2%Rank03Participants3Percentile0.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkIMO-AnswerBenchScore90.0%Rank03Participants20Percentile89.5%EvidenceCEvaluatedAug 17, 2026
BenchmarkSkillsBenchScore55.3%Rank03Participants8Percentile71.4%EvidenceCEvaluatedAug 17, 2026
BenchmarkSuperChemScore54.9%Rank03Participants3Percentile0.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkUSAMO 2026Score30.24 pointsRank03Participants3Percentile0.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkWildClawBenchScore53.6%Rank03Participants5Percentile50.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkClaw-EvalScore68.5%Rank04Participants14Percentile76.9%EvidenceCEvaluatedAug 17, 2026
BenchmarkDeepSearchQAScore91.0%Rank04Participants9Percentile62.5%EvidenceCEvaluatedAug 17, 2026
BenchmarkFrontierScience ResearchScore21.3%Rank04Participants4Percentile0.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkWideSearchScore76.4%Rank04Participants10Percentile66.7%EvidenceCEvaluatedAug 17, 2026
BenchmarkMathArena ApexScore38.7%Rank05Participants8Percentile42.9%EvidenceCEvaluatedAug 17, 2026
BenchmarkAPEX-AgentsScore25.6%Rank08Participants8Percentile0.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkMCP AtlasScore79.1%Rank08Participants33Percentile78.1%EvidenceCEvaluatedAug 17, 2026
BenchmarkNL2RepoScore45.6%Rank08Participants17Percentile56.3%EvidenceCEvaluatedAug 17, 2026
BenchmarkDeepSWEScore28.0%Rank10Participants11Percentile10.0%EvidenceCEvaluatedAug 17, 2026
BenchmarkSWE-bench MultilingualScore75.8%Rank10Participants38Percentile75.7%EvidenceCEvaluatedAug 17, 2026
BenchmarkBrowseCompScore84.2%Rank14Participants62Percentile78.7%EvidenceCEvaluatedAug 17, 2026
BenchmarkTerminal-Bench 2.1Score71.7%Rank17Participants28Percentile40.7%EvidenceCEvaluatedAug 17, 2026
BenchmarkGPQAScore90.4%Rank21Participants239Percentile91.6%EvidenceCEvaluatedAug 17, 2026
BenchmarkSWE-Bench ProScore57.9%Rank21Participants50Percentile59.2%EvidenceCEvaluatedAug 17, 2026
BenchmarkSWE-Bench VerifiedScore78.0%Rank22Participants111Percentile80.9%EvidenceCEvaluatedAug 17, 2026
BenchmarkToolathlonScore48.5%Rank22Participants37Percentile41.7%EvidenceCEvaluatedAug 17, 2026

Hy3 Arena Results

Preference and agent-evaluation results for the default version.

14 rows
Columns

Show columns

Sort by
Arena
Category
Rank
Rating / score
Votes
Observations
Result date
ArenawebdevCategoryoverallRank24Rating / score1521.7Votes2,284ObservationsN/AResult dateAug 14, 2026
Arenaagent praise complaintCategoryoverallRank26Rating / score0.0VotesN/AObservations2.6KResult dateAug 13, 2026
Arenaagent bash recovery stepsCategoryoverallRank29Rating / score0.0VotesN/AObservations16KResult dateAug 13, 2026
ArenaagentCategoryoverallRank32Rating / score-0.0VotesN/AObservations620.8KResult dateAug 13, 2026
Arenaagent task outcome explicitCategoryoverallRank36Rating / score-0.0VotesN/AObservations8.3KResult dateAug 13, 2026
Arenaagent steerabilityCategoryoverallRank42Rating / score-0.1VotesN/AObservations10.3KResult dateAug 13, 2026
Arenaagent tool hallucinationCategoryoverallRank45Rating / score-0.0VotesN/AObservations583.7KResult dateAug 13, 2026
Arenatext style controlCategoryoverallRank53Rating / score1456.9Votes4,650ObservationsN/AResult dateAug 12, 2026
ArenatextCategoryoverallRank57Rating / score1441.5Votes4,650ObservationsN/AResult dateAug 12, 2026
Arenatext factualityCategoryoverallRank57Rating / score1448.2Votes4,650ObservationsN/AResult dateAug 12, 2026
ArenawebdevCategoryoverallRank83Rating / score1356.1Votes1,394ObservationsN/AResult dateAug 14, 2026
Arenatext factualityCategoryoverallRank117Rating / score1415.8Votes6,621ObservationsN/AResult dateAug 12, 2026
Arenatext style controlCategoryoverallRank123Rating / score1412.8Votes6,624ObservationsN/AResult dateAug 12, 2026
ArenatextCategoryoverallRank127Rating / score1405.1Votes6,624ObservationsN/AResult dateAug 12, 2026

Hy3 Pricing

Official vendor API pricing appears first, followed by individual provider offers.

Official API
N/A
Official provider
N/A
Lowest third-party
From $0.066 input, $0.26 output per 1M via NanoGPT
Tracked offerings
12
12 rows
Columns

Show columns

Sort by
Provider
Provider model ID
Region
Input / 1M
Output / 1M
Context
Updated
ProviderTencent TokenHubProvider model IDhy3RegionglobalInput / 1MN/AOutput / 1MN/AContext256KUpdatedAug 17, 2026
ProviderTencent Token PlanProvider model IDhy3RegionglobalInput / 1MN/AOutput / 1MN/AContext256KUpdatedAug 17, 2026
ProviderNanoGPTProvider model IDtencent/hy3RegionglobalInput / 1M$0.066Output / 1M$0.26Context262.1KUpdatedAug 17, 2026
ProviderOpenRouterProvider model IDtencent/hy3RegionglobalInput / 1M$0.132Output / 1M$0.528Context262.1KUpdatedAug 17, 2026
ProviderDeep InfraProvider model IDtencent/Hy3RegionglobalInput / 1M$0.14Output / 1M$0.58Context262.1KUpdatedAug 17, 2026
ProviderHugging FaceProvider model IDtencent/Hy3RegionglobalInput / 1M$0.14Output / 1M$0.58Context262.1KUpdatedAug 17, 2026
ProviderOpenCode GoProvider model IDhy3RegionglobalInput / 1M$0.14Output / 1M$0.58Context256KUpdatedAug 17, 2026
ProviderVercel AI GatewayProvider model IDtencent/hy3RegionglobalInput / 1M$0.14Output / 1M$0.58Context262.1KUpdatedAug 17, 2026
ProviderEden AIProvider model IDdeepinfra/tencent/Hy3RegionglobalInput / 1M$0.14Output / 1M$0.58Context262.1KUpdatedAug 17, 2026
ProviderLLM GatewayProvider model IDhy3RegionglobalInput / 1M$0.14Output / 1M$0.58Context262.1KUpdatedAug 17, 2026
ProviderKilo GatewayProvider model IDtencent/hy3RegionglobalInput / 1M$0.14Output / 1M$0.58Context262.1KUpdatedAug 17, 2026
ProviderCrossModelProvider model IDtencent/hy3RegionglobalInput / 1M$0.16Output / 1M$0.64Context262.1KUpdatedAug 17, 2026

Official prices use only the vendor's configured official Provider and positive standard USD PAYG rates. Third-party offers remain explicitly labeled.

Hy3 Runtime Performance

Provider-specific output speed and catalog latency for Hy3. Runtime does not affect the capability score.

No runtime data

No provider-specific speed or latency record is linked to the default version yet.

Browse runtime rankings

Output Speed is generated output tokens received per second. Catalog latency is reported separately from observed provider TTFT.

Hy3 Specifications

Technical details for the model's default version.

Version
Hy3
Released
Jul 6, 2026
Knowledge cutoff
Unknown
Parameters
295B
Context window
256K
Max output
64K
Inputs
text
Outputs
text
Open weights
Yes
License
Apache 2.0

Hy3 Versions

Available versions of this model. The score column identifies the version used in the overall ranking.

1 rows
Columns

Show columns

Sort by
Version
Released
LLMBoard
Parameters
Context
Max output
Open weights
License
VersionHy3ReleasedJul 6, 2026LLMBoard70.7Parameters295BContext256KMax output64KOpen weightsYesLicenseApache 2.0

Hy3 vs nearby models

Open a comparison with the three ranked models immediately above and below this model.

Hy3vsSakana NamazuHy3vsGPT-5.2-ProHy3vsMiniMax M3Hy3vsGPT-5.2Hy3vsLaguna S 2.1Hy3vsClaude Opus 4.5

Models similar to Hy3

Recommendations prioritize the same model type and family, then the closest LLMBoard score.

#38+0.4
MI

MiniMax M3

MiniMax

71.1 LLMBoard

DetailsCompare
#40-0.6
OP

GPT-5.2

OpenAI

70.0 LLMBoard

DetailsCompare
#41-0.9
PO

Laguna S 2.1

Poolside

69.8 LLMBoard

DetailsCompare
#42-1.4
AN

Claude Opus 4.5

Anthropic

69.3 LLMBoard

DetailsCompare
#43-1.4
ZA

GLM 5.1

Zhipu AI

69.3 LLMBoard

DetailsCompare
#37+1.4
OP

GPT-5.2-Pro

OpenAI

72.1 LLMBoard

DetailsCompare

What is Hy3?

Key information about Hy3 and its available data.

8B MTP layer, developed by the Tencent Hy Team. Following the Hy3 Preview launch in late April, the team scaled up post-training with higher-quality data and RL, gathering feedback from 50+ products.

Hy3 outperforms similar-size models and rivals flagship open-source models with 2-5x the parameters, with strong gains in reasoning, agentic, and long-context tasks.

It uses 80 layers (plus 1 MTP layer), 64 GQA attention heads (8 KV heads, head dim 128), a 4096 hidden size, 192 experts with top-8 activated, a 256K context window, and BF16 precision.

Hy3 is a hybrid-thinking model supporting configurable reasoning effort (no_think, low, high), and emphasizes production-grade tool-call and output-format stability, reduced hallucination, and reliable multi-turn intent tracking.

Data as of 2026-08-17.

FAQ

Common questions about Hy3.

When was Hy3 released?

Hy3's default version was released on Jul 6, 2026.

How much does Hy3 cost?

No official standard PAYG price is currently available for Hy3. The lowest tracked third-party offer starts at $0.066 input and $0.26 output via NanoGPT.

Who created Hy3?

Hy3 was created by Tencent.

What is the context window for Hy3?

The default version has a 256K token context window.

Is Hy3 open weight?

Yes. The default version is marked as open weight under Apache 2.0.

How many API providers offer Hy3?

12 provider offerings are linked to the default version.

What models should I compare Hy3 with?

Nearby ranked alternatives include Sakana Namazu, GPT-5.2-Pro, MiniMax M3.