Passer au contenu principal

Llama 3.2 Instruct 90B (Vision)

MetaLlamaOpen WeightLlama 3.2 · Usage Commercial

Description

Llama 3.2 90B is a large multimodal language model optimized for visual recognition, image reasoning, and captioning tasks. It supports a context length of 128,000 tokens and is designed for deployment on edge and mobile devices, offering state-of-the-art performance in image understanding and generative tasks.

Date de sortie
2024-09-25
Paramètres
90.0B
Longueur du contexte
—
Modalités
image, text

Radar de capacités

24
general
21
coding
30
reasoning
31
science
27
agents
85
multimodal

Classements

Domaine#RangScoreSource
Classement codage483
23.0
AA
Classement général505
28.0
AA
Classement multimodal105
43.0
LS
Science536
23.0
AA

Scores de benchmarks (LLM Stats)

(LLM Stats (zeroeval))

General

MMLU86.0%Aut.

Math

MGSM86.9%Aut.
MATH68.0%Aut.
MathVista57.3%Aut.

Multimodal

MMMU60.3%Aut.

Reasoning

ChartQAMasry et al. (2022)85.5%Aut.
GPQANYU + Cohere + Anthropic (2023)46.7%Aut.

Vision

AI2D92.3%Aut.
DocVQADocVQA (2020)90.1%Aut.
VQAv278.1%Aut.
TextVQA73.5%Aut.
InfographicsQA56.8%Aut.
MMMU-Pro45.2%Aut.

Indices d'évaluation AA

(Artificial Analysis)
Mmlu Pro(TIGER-Lab (Univ. of Waterloo, Toronto, CMU, 2024))
67.1
Math 500(OpenAI (2024), subset of Hendrycks et al. MATH (2021))
62.9
Gpqa(NYU + Cohere + Anthropic (2023))
43.2
Livecodebench(UC Berkeley + MIT + Cornell (2024))
21.4
Intelligence Index(Artificial Analysis)
6.4
Aime(MAA (Mathematical Association of America))
5.0
Hle(Center for AI Safety + Scale AI (2025))
4.5

Scores par catégorie LLM Stats

(LLM Stats (zeroeval))
Language
90
Legal
90
Finance
90
Image To Text
80
Math
70
Multimodal
70
Reasoning
70
General
70
Healthcare
70
Vision
70
Physics
50
Biology
50
Chemistry
50

Tarification

Prix d'entréeGratuit
Prix de sortieGratuit
Prix mixte (3:1)Gratuit

Vitesse

Tokens/sec0.0
Délai du premier token0.00s
Temps de réponse0.00s

Classement des Prix par Fournisseur

Aucune donnée de fournisseur disponible

Sources externes