Tencent model product
8B MTP layer, developed by the Tencent Hy Team.
Updated Aug 12, 2026. Default version: Hy3
Technical details for the model's default version.
This profile uses the model's current scored version. Arena ratings and prices are shown separately.
Benchmark scores for Hy3.
| CL-bench | 23.8% | 01 | 2 | 100.0% | C | |
| CL-bench (Life) | 17.0% | 01 | 1 | 100.0% | C | |
| CMT-Benchmark | 37.9% | 01 | 1 | 100.0% | C | |
| HorizonMath | 7.1% | 01 | 3 | 100.0% | C | |
| Humanity's Last Exam (no tools, text-only) | 47.0% | 01 | 1 | 100.0% | C | |
| PHYBench | 77.4% | 01 | 1 | 100.0% | C | |
| AA-LCR | 73.4% | 02 | 16 | 93.3% | C | |
| ArXivMath | 52.2% | 02 | 2 | 0.0% | C | |
| Humanity's Last Exam (with tools, text-only) | 53.2% | 02 | 2 | 0.0% | C | |
| FrontierScience Olympiad | 74.8% | 03 | 3 | 0.0% | C | |
| IMO-AnswerBench | 90.0% | 03 | 19 | 88.9% | C | |
| SkillsBench | 55.3% | 03 | 8 | 71.4% | C | |
| SuperChem | 54.9% | 03 | 3 | 0.0% | C | |
| USAMO 2026 | 30.24 points | 03 | 3 | 0.0% | C | |
| WideSearch | 76.4% | 03 | 9 | 75.0% | C | |
| WildClawBench | 53.6% | 03 | 5 | 50.0% | C | |
| Claw-Eval | 68.5% | 04 | 13 | 75.0% | C | |
| DeepSearchQA | 91.0% | 04 | 9 | 62.5% | C | |
| FrontierScience Research | 21.3% | 04 | 4 | 0.0% | C | |
| MathArena Apex | 38.7% | 04 | 7 | 50.0% | C | |
| NL2Repo | 45.6% | 06 | 14 | 61.5% | C | |
| APEX-Agents | 25.6% | 07 | 7 | 0.0% | C | |
| MCP Atlas | 79.1% | 07 | 31 | 80.0% | C | |
| DeepSWE | 28.0% | 09 | 10 | 11.1% | C | |
| SWE-bench Multilingual | 75.8% | 09 | 34 | 75.8% | C | |
| Terminal-Bench 2.1 | 71.7% | 13 | 19 | 33.3% | C | |
| BrowseComp | 84.2% | 14 | 58 | 77.2% | C | |
| Toolathlon | 48.5% | 17 | 31 | 46.7% | C | |
| SWE-Bench Pro | 57.9% | 19 | 45 | 59.1% | C | |
| SWE-Bench Verified | 78.0% | 20 | 105 | 81.7% | C | |
| GPQA | 90.4% | 21 | 234 | 91.4% | C |
Preference and agent-evaluation results for the default version.
| agent praise complaint | overall | 19 | 0.0 | N/A | 2K | |
| webdev | overall | 20 | 1524.2 | 2,175 | N/A | |
| agent bash recovery steps | overall | 28 | 0.0 | N/A | 12.6K | |
| agent | overall | 30 | -0.0 | N/A | 477.6K | |
| agent task outcome explicit | overall | 31 | -0.0 | N/A | 6.7K | |
| agent steerability | overall | 39 | -0.1 | N/A | 8.1K | |
| agent tool hallucination | overall | 41 | -0.0 | N/A | 448.2K | |
| text style control | overall | 52 | 1456.3 | 4,544 | N/A | |
| text | overall | 58 | 1441.0 | 4,544 | N/A | |
| text factuality | overall | 59 | 1448.0 | 4,544 | N/A | |
| webdev | overall | 79 | 1356.3 | 1,394 | N/A | |
| text factuality | overall | 117 | 1415.7 | 6,624 | N/A | |
| text style control | overall | 122 | 1412.8 | 6,627 | N/A | |
| text | overall | 127 | 1405.0 | 6,627 | N/A |
Provider-specific output speed and catalog latency for Hy3. Runtime does not affect the capability score.
No provider-specific speed or latency record is linked to the default version yet.
Output Speed is generated output tokens received per second. Catalog latency is reported separately from observed provider TTFT.
Official vendor API pricing appears first, followed by individual provider offers.
| Tencent TokenHub | hy3 | global | N/A | N/A | 256K | |
| Tencent Token Plan | hy3 | global | N/A | N/A | 256K | |
| NanoGPT | tencent/hy3 | global | $0.066 | $0.26 | 262.1K | |
| OpenRouter | tencent/hy3 | global | $0.132 | $0.528 | 262.1K | |
| OpenCode Go | hy3 | global | $0.14 | $0.58 | 256K | |
| LLM Gateway | hy3 | global | $0.14 | $0.58 | 262.1K | |
| Kilo Gateway | tencent/hy3 | global | $0.14 | $0.58 | 262.1K | |
| Vercel AI Gateway | tencent/hy3 | global | $0.14 | $0.58 | 262.1K | |
| Hugging Face | tencent/Hy3 | global | $0.14 | $0.58 | 262.1K | |
| Deep Infra | tencent/Hy3 | global | $0.14 | $0.58 | 262.1K | |
| CrossModel | tencent/hy3 | global | $0.16 | $0.64 | 262.1K |
Official prices use only the vendor's configured official Provider and positive standard USD PAYG rates. Third-party offers remain explicitly labeled.
Available versions of this model. The score column identifies the version used in the overall ranking.
| Hy3 | 71.0 | 295B | 256K | 64K | Yes | Apache 2.0 |
Key information about Hy3 and its available data.
8B MTP layer, developed by the Tencent Hy Team. Following the Hy3 Preview launch in late April, the team scaled up post-training with higher-quality data and RL, gathering feedback from 50+ products.
Hy3 outperforms similar-size models and rivals flagship open-source models with 2-5x the parameters, with strong gains in reasoning, agentic, and long-context tasks.
It uses 80 layers (plus 1 MTP layer), 64 GQA attention heads (8 KV heads, head dim 128), a 4096 hidden size, 192 experts with top-8 activated, a 256K context window, and BF16 precision.
Hy3 is a hybrid-thinking model supporting configurable reasoning effort (no_think, low, high), and emphasizes production-grade tool-call and output-format stability, reduced hallucination, and reliable multi-turn intent tracking.
Data as of 2026-08-11.
Open a comparison with the three ranked models immediately above and below this model.
Recommendations prioritize the same model type and family, then the closest LLMBoard score.
Common questions about Hy3.
Hy3's default version was released on Jul 6, 2026.
No official standard PAYG price is currently available for Hy3. The lowest tracked third-party offer starts at $0.066 input and $0.26 output via NanoGPT.
Hy3 was created by Tencent.
The default version has a 256K token context window.
Yes. The default version is marked as open weight under Apache 2.0.
11 provider offerings are linked to the default version.
Nearby ranked alternatives include GPT-5.2-Pro, GPT-5.2, MiniMax M3.