跳转到主要内容

Hy3

Tencent开源权重Apache 2.0 · 商用许可

描述

Hy3 is a 295B-parameter Mixture-of-Experts (MoE) model with 21B active parameters and a 3.8B MTP layer, developed by the Tencent Hy Team. Following the Hy3 Preview launch in late April, the team scaled up post-training with higher-quality data and RL, gathering feedback from 50+ products. Hy3 outperforms similar-size models and rivals flagship open-source models with 2-5x the parameters, with strong gains in reasoning, agentic, and long-context tasks. It uses 80 layers (plus 1 MTP layer), 64 GQA attention heads (8 KV heads, head dim 128), a 4096 hidden size, 192 experts with top-8 activated, a 256K context window, and BF16 precision. Hy3 is a hybrid-thinking model supporting configurable reasoning effort (no_think, low, high), and emphasizes production-grade tool-call and output-format stability, reduced hallucination, and reliable multi-turn intent tracking.

发布日期
2026-07-06
参数规模
295.0B
上下文长度
262K
支持模态
text

能力雷达图

27
general
57
coding
90
reasoning
64
science
70
agents
0
multimodal

排行榜排名

领域#排名分数来源
智能体能力模型榜53
48.0
LS
代码能力榜128
78.0
AA
通用能力榜321
41.0
AA
科学能力111
72.0
AA

基准测试分数 (LLM Stats)

(LLM Stats (zeroeval))

Agents

WildClawBench53.6%自报
Toolathlon48.5%自报

Chemistry

SuperChem54.9%自报

Code

Claw-Eval68.5%自报
SkillsBench55.3%自报
NL2Repo45.6%自报
DeepSWE28.0%自报
CL-bench23.8%自报
CL-bench (Life)17.0%自报

Math

USAMO 202630.24 / 42自报
IMO-AnswerBench90.0%自报
ArXivMath52.2%自报
MathArena Apex38.7%自报
HorizonMath7.1%自报

Physics

PHYBench77.4%自报
CMT-Benchmark37.9%自报

Reasoning

GPQANYU + Cohere + Anthropic (2023)90.4%自报
BrowseCompOpenAI (2025)84.2%自报
MCP Atlas79.1%自报
SWE-Bench Verified78.0%自报
SWE-bench Multilingual75.8%自报
FrontierScience Olympiad74.8%自报
AA-LCR73.4%自报
Terminal-Bench 2.171.7%自报
SWE-Bench ProPrinceton NLP (2024)57.9%自报
Humanity's Last Exam (with tools, text-only)53.2%自报
Humanity's Last Exam (no tools, text-only)47.0%自报
APEX-Agents25.6%自报
FrontierScience Research21.3%自报

Search

DeepSearchQA91.0%自报
WideSearch76.4%自报

AA 评测指数

(Artificial Analysis)
Gpqa(NYU + Cohere + Anthropic (2023))
89.7
Lcr(Artificial Analysis)
79.0
Terminalbench V2 1
64.4
Coding Index(Artificial Analysis)
58.8
Scicode(UIUC + Argonne National Lab (2024))
48.6
Hle(Center for AI Safety + Scale AI (2025))
33.5
Intelligence Index(Artificial Analysis)
25.3
Tau Banking
22.9
Terminalbench V4 0
0.5

LLM Stats 分类评分

(LLM Stats (zeroeval))
Math
4
Reasoning
2
General
2
Biology
90
Physics
80
Search
80
Frontend Development
80
Long Context
70
Chemistry
70
Tool Calling
70
Science
60
Agents
60
Code
60
Knowledge
50
Coding
50

定价

输入价格$0.136 / 1M tokens
输出价格$0.555 / 1M tokens
混合价格(3:1)$0.241 / 1M tokens
缓存读取价格$0.033 / 1M tokens

速度

Tokens/秒87.7
首Token延迟2.06s
首回答延迟24.85s

供应商价格排行

供应商价格排行

14 个供应商

最便宜: DeepInfra最贵: OrcaRouter
供应商输入输出
1DeepInfra最便宜
$0
$0
2NanoGPT
$0.066
$0.26
3Kilo Gateway
$0.13
$0.53
4OpenRouter
$0.132
$0.528
5DevPass (LLM Gateway)
$0.132
$0.528
6LLM Gateway
$0.132
$0.528
7Tencent主要
$0.136
$0.555
8OpenCode Go
$0.14
$0.58
9Requesty
$0.14
$0.58
10Vercel AI Gateway
$0.14
$0.58
11Jalapeno Cloud
$0.14
$0.58
12AIHubMix
$0.1562
$0.6248
13CrossModel
$0.16
$0.64
14OrcaRouter
$0.18
$0.59

比较该模型在不同 API 供应商之间的定价。

外部链接