メインコンテンツへスキップ

Llama 3.3 Instruct 70B

MetaLlamaオープンウエイトLlama 3.3 Community License Agreement

説明

Llama 3.3 is a multilingual large language model optimized for dialogue use cases across multiple languages. It is a pretrained and instruction-tuned generative model with 70 billion parameters, outperforming many open-source and closed chat models on common industry benchmarks. Llama 3.3 supports a context length of 128,000 tokens and is designed for commercial and research use in multiple languages.

リリース日
2024-12-06
パラメータ
70.0B
コンテキスト長
131K
モダリティ
text

能力レーダー

26
general
18
coding
28
reasoning
36
science
80
agents
0
multimodal

ランキング

ドメイン#順位スコアソース
コーディングランキング532
17.0
AA
総合ランキング409
34.0
AA
科学519
25.0
AA

ベンチマークスコア (LLM Stats)

(LLM Stats (zeroeval))

Chat

IFEvalGoogle Research (2023)92.1%自己申告

General

MMLU86.0%自己申告
BFCL v277.3%自己申告

Language

MMLU-Pro68.9%自己申告

Math

MGSM91.1%自己申告
MATH77.0%自己申告

Reasoning

HumanEvalOpenAI (2021)88.4%自己申告
MBPP EvalPlus87.6%自己申告
GPQANYU + Cohere + Anthropic (2023)50.5%自己申告

AA評価指数

(Artificial Analysis)
Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))
77.3
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))
71.3
Gpqa(NYU + Cohere + Anthropic (2023))
49.8
Ifbench(Google Research (2023))
47.1
Aime(MAA (Mathematical Association of America))
30.0
Livecodebench(UC Berkeley + MIT + Cornell (2024))
28.8
Tau2(Sierra + U Toronto + Vector Institute (2025))
26.6
Lcr(Artificial Analysis)
15.7
Coding Index(Artificial Analysis)
11.9
Intelligence Index(Artificial Analysis)
7.7
Math Index(Artificial Analysis)
7.7
Aime 25(MAA (Mathematical Association of America))
7.7
Terminalbench V2 1
4.9
Hle(Center for AI Safety + Scale AI (2025))
3.6
Terminalbench Hard(Stanford × Laude Institute (2026))
3.0

LLM Statsカテゴリスコア

(LLM Stats (zeroeval))
Chat
90
Instruction Following
90
Structured Output
90
Code
90
Language
80
Legal
80
Math
80
Reasoning
80
Finance
80
General
80
Healthcare
80
Tool Calling
80
Physics
50
Biology
50
Chemistry
50

価格設定

入力価格$0.71 / 1Mトークン
出力価格$0.72 / 1Mトークン
混合価格(3:1)$0.712 / 1Mトークン

速度

トークン/秒90.7
初トークン遅延0.64s
初回答遅延0.64s

プロバイダー価格ランキング

プロバイダー価格ランキング

6 プロバイダー

最安: DeepInfra最高: Meta
プロバイダー入力出力
1DeepInfra最安
$0
$0
2NanoGPT
$0.05
$0.23
3OpenRouter
$0.1
$0.32
4Kilo Gateway
$0.1
$0.32
5NovitaAI
$0.135
$0.4
6Metaプライマリ
$0.71
$0.72

このモデルの異なるAPIプロバイダー間の価格を比較。

外部リンク