Mercury 2
InceptionProprietary
설명
Mercury 2 is the fastest reasoning LLM, built on diffusion-based language model (dLLM) architecture. Instead of generating text token-by-token, it refines multiple text blocks simultaneously, achieving over 1,000 tokens per second on Nvidia Blackwell GPUs — 5x faster than leading speed-optimized LLMs. Supports tool usage and JSON output with 128K context window.
출시일
2026-02-20
파라미터
—
컨텍스트 길이
128K
모달리티
text
능력 레이더
21
general
32
coding
77
reasoning
52
science
50
agents
0
multimodal
랭킹
벤치마크 점수 (LLM Stats)
(LLM Stats (zeroeval))Biology
GPQANYU + Cohere + Anthropic (2023)
74.0%자체 보고
SciCode
38.0%자체 보고
Code
LiveCodeBench
67.0%자체 보고
Communication
Tau2 Airline
53.0%자체 보고
General
IFBench
71.0%자체 보고
Math
AIME 2025
91.1%자체 보고
AA 평가 지수
(Artificial Analysis)Coding Index(Artificial Analysis)31.1
Intelligence Index(Artificial Analysis)21.9
Gpqa(NYU + Cohere + Anthropic (2023))0.8
Tau2(Sierra + U Toronto + Vector Institute (2025))0.7
Ifbench(Google Research (2023))0.7
Lcr(Artificial Analysis)0.4
Scicode(UIUC + Argonne National Lab (2024))0.4
Terminalbench V2 10.3
Terminalbench Hard(Stanford × Laude Institute (2026))0.3
Hle(Center for AI Safety + Scale AI (2025))0.2
Tau Banking0.1
LLM Stats 카테고리 점수
(LLM Stats (zeroeval))Instruction Following70
General70
Math60
Physics60
Reasoning60
Biology60
Chemistry60
Code50
Communication50
Tool Calling50
가격
입력 가격$0.25 / 1M 토큰
출력 가격$0.75 / 1M 토큰
혼합 가격 (3:1)$0.375 / 1M 토큰
캐시 읽기 가격$0.025 / 1M 토큰
속도
토큰/초912.9
첫 토큰 지연3.96s
첫 응답 지연3.96s
공급자 가격 순위
공급자 가격 순위
6개 공급자
최저가: Inception최고가: Venice AI
공급자입력출력
1Inception최저가
$0
$0
2NanoGPT
$0.25
$0.75
3OpenRouter
$0.25
$0.75
4Kilo Gateway
$0.25
$0.75
5Vercel AI Gateway
$0.25
$0.75
6Venice AI
$0.3125
$0.9375
이 모델의 다양한 API 공급자 간 가격 비교.