Skip to main content

Hy3

TencentOpen WeightApache 2.0 · Commercial OK

Description

Hy3 is a 295B-parameter Mixture-of-Experts (MoE) model with 21B active parameters and a 3.8B MTP layer, developed by the Tencent Hy Team. Following the Hy3 Preview launch in late April, the team scaled up post-training with higher-quality data and RL, gathering feedback from 50+ products. Hy3 outperforms similar-size models and rivals flagship open-source models with 2-5x the parameters, with strong gains in reasoning, agentic, and long-context tasks. It uses 80 layers (plus 1 MTP layer), 64 GQA attention heads (8 KV heads, head dim 128), a 4096 hidden size, 192 experts with top-8 activated, a 256K context window, and BF16 precision. Hy3 is a hybrid-thinking model supporting configurable reasoning effort (no_think, low, high), and emphasizes production-grade tool-call and output-format stability, reduced hallucination, and reliable multi-turn intent tracking.

Release Date
2026-07-06
Parameters
295.0B
Context Length
262K
Modalities
text

Capability Radar

40
general
57
coding
90
reasoning
64
science
70
agents
0
multimodal

Rankings

Domain#RankScoreSource
Agentic Capability51
52.0
LS
Code Ranking57
80.0
AA
General Ranking128
66.0
AA
Science61
78.0
AA

Benchmark Scores (LLM Stats)

(LLM Stats (zeroeval))

Agents

DeepSearchQA91.0%SR
BrowseCompOpenAI (2025)84.2%SR
MCP Atlas79.1%SR
WideSearch76.4%SR
Terminal-Bench 2.171.7%SR
Claw-Eval68.5%SR
SWE-Bench ProPrinceton NLP (2024)57.9%SR
SkillsBench55.3%SR
WildClawBench53.6%SR
Toolathlon48.5%SR
NL2Repo45.6%SR
DeepSWE28.0%SR
APEX-Agents25.6%SR
CL-bench23.8%SR
CL-bench (Life)17.0%SR

Biology

GPQANYU + Cohere + Anthropic (2023)90.4%SR

Chemistry

SuperChem54.9%SR

Code

SWE-Bench Verified78.0%SR
SWE-bench Multilingual75.8%SR

Long Context

AA-LCR73.4%SR

Math

USAMO 202630.24 / 42SR
IMO-AnswerBench90.0%SR
FrontierScience Olympiad74.8%SR
Humanity's Last Exam (with tools, text-only)53.2%SR
ArXivMath52.2%SR
Humanity's Last Exam (no tools, text-only)47.0%SR
MathArena Apex38.7%SR
HorizonMath7.1%SR

Physics

PHYBench77.4%SR
CMT-Benchmark37.9%SR

Reasoning

FrontierScience Research21.3%SR

AA Evaluation Indices

(Artificial Analysis)
Coding Index(Artificial Analysis)
58.8
Intelligence Index(Artificial Analysis)
42.2
Gpqa(NYU + Cohere + Anthropic (2023))
0.9
Lcr(Artificial Analysis)
0.7
Terminalbench V2 1
0.6
Scicode(UIUC + Argonne National Lab (2024))
0.5
Hle(Center for AI Safety + Scale AI (2025))
0.3
Tau Banking
0.2

LLM Stats Category Scores

(LLM Stats (zeroeval))
Math
6
Reasoning
2
General
2
Physics
90
Biology
90
Search
80
Frontend Development
80
Long Context
70
Chemistry
70
Tool Calling
70
Agents
60
Code
60
Science
50
Knowledge
50
Coding
50

Pricing

Input Price$0.136 / 1M tokens
Output Price$0.557 / 1M tokens
Blended Price (3:1)$0.241 / 1M tokens
Cache Read Price$0.033 / 1M tokens

Speed

Tokens/sec68.0
Time to First Token1.81s
Time to Answer31.22s

Provider Price Ranking

Provider Price Ranking

8 providers

Cheapest: NanoGPTMost Expensive: CrossModel
ProviderInputOutput
1NanoGPTCheapest
$0.066
$0.26
2OpenRouter
$0.132
$0.528
3TencentPRIMARY
$0.136
$0.557
4OpenCode Go
$0.14
$0.58
5Kilo Gateway
$0.14
$0.58
6Vercel AI Gateway
$0.14
$0.58
7LLM Gateway
$0.14
$0.58
8CrossModel
$0.16
$0.64

Compare pricing across different API providers for this model.

External Sources