Seed 1.8
ByteDanceProprietary
Description
Optimized specifically for multimodal agent scenarios. It features enhanced agent capabilities, upgraded multimodal comprehension, and more flexible context management.
Release Date
2026-02-17
Parameters
—
Context Length
—
Modalities
image, text, video
Capability Radar
39
general
100
coding
70
reasoning
51
scienceest.
50
agents
70
multimodal
Science is estimated from LLM Stats science scores or reasoning when no dedicated science benchmarks are available.
Rankings
| Domain | #Rank | Score | Source |
|---|---|---|---|
| Agentic Capability | 85 | 37.0 | LS |
| Multimodal Ranking | 91 | 48.0 | LS |
Benchmark Scores (LLM Stats)
(LLM Stats (zeroeval))Agents
AndroidWorld
70.7%SR
Chat
Multi-Challenge
66.7%SR
Coding
AetherCode
38.2%SR
General
MMLU
92.3%SR
Language
MMLU-Pro
84.9%SR
Math
AIME 2025
94.3%SR
MathVista
87.7%SR
MathVision
81.3%SR
Beyond AIME
77.0%SR
DynaMath
61.5%SR
AMO Bench
60.0%SR
Multimodal
MMMU
83.4%SR
VideoMMMU
82.7%SR
OSWorld
61.9%SR
Physics
PHYBench
41.0%SR
Reasoning
LiveCodeBench Pro
1930.00 / 3000SR
BrowseComp-zh
81.3%SR
LiveCodeBench v6
79.5%SR
SWE-Bench Verified
72.9%SR
CharXiv-R
71.4%SR
ARC-AGI
67.9%SR
BrowseCompOpenAI (2025)
67.6%SR
SuperGPQA
64.8%SR
Terminal-Bench 2.0Stanford × Laude Institute (2026)
45.2%SR
Multi-SWE-Bench
42.0%SR
Humanity's Last Exam (with tools, text-only)
40.9%SR
Search
WideSearch
63.8%SR
Tool Calling
BFCL-V4
57.2%SR
Video
LiveSports-3K
77.5%SR
OVOBench
72.6%SR
VideoSimpleQA
67.8%SR
OVBench
65.1%SR
Minerva
62.4%SR
TOMATO
60.8%SR
Vision
CountBench
96.30 / 100SR
SimpleVQA
65.40 / 100SR
RefSpatialBench
56.30 / 100SR
ZEROBench
11.00 / 100SR
VLMsAreBlind
93.0%SR
AI2D
89.1%SR
VideoMME w sub.
87.8%SR
TempCompass
86.9%SR
MMStar
79.9%SR
MuirBench
78.7%SR
LongVideoBench
77.4%SR
BLINK
74.3%SR
MMMU-Pro
73.2%SR
MMVU
73.1%SR
LVBench
73.0%SR
TVBench
71.5%SR
MotionBench
70.6%SR
DUDE
69.4%SR
Hallusion Bench
63.9%SR
VLMsAreBiased
62.0%SR
EMMA
60.9%SR
ERQA
58.8%SR
OmniDocBench 1.5
10.6%SR
AA Evaluation Indices
(Artificial Analysis)No AA evaluation data available
LLM Stats Category Scores
(LLM Stats (zeroeval))Code100
Image To Text65
Grounding56
Reasoning53
General39
Spatial Reasoning31
Vision7
Multimodal3
Language90
Legal80
Finance80
Healthcare80
Chat70
Knowledge70
Long Context70
Math70
Search70
Frontend Development70
3d70
Communication70
Video70
Agents60
Chemistry60
Economics60
Physics50
Tool Calling50
Science40
Coding40
Structured Output10
Pricing
Input Price$0 / 1M tokens
Output Price$0 / 1M tokens
Blended Price (3:1)$0 / 1M tokens
Speed
No speed data available
Provider Price Ranking
Provider Price Ranking
3 providers
Cheapest: ByteDanceMost Expensive: Requesty
ProviderInputOutput
1ByteDancePRIMARY
$0
$0
2DeepInfra
$0
$0
3Requesty
$0.25
$2
Compare pricing across different API providers for this model.