Qwen3 Next 80B A3B Instruct
説明
Qwen3-Next-80B-A3B-Instruct is the first in the Qwen3-Next series, featuring groundbreaking architectural innovations. It uses Hybrid Attention combining Gated DeltaNet and Gated Attention for efficient ultra-long context modeling, High-Sparsity MoE with 512 experts (10 activated + 1 shared) achieving extreme low activation ratio, and Multi-Token Prediction for improved performance and faster inference. With 80B total parameters and only 3B activated, it outperforms Qwen3-32B-Base with 10% training cost and 10x throughput for 32K+ contexts. The model performs on par with Qwen3-235B-A22B-Instruct-2507 while excelling at ultra-long-context tasks up to 256K tokens (extensible to 1M with YaRN). Architecture: 48 layers, 15T training tokens, hybrid layout of 12*(3*(Gated DeltaNet->MoE)->(Gated Attention->MoE)).
能力レーダー
ランキング
| ドメイン | #順位 | スコア | ソース |
|---|---|---|---|
| エージェント能力 | 66 | 48.0 | LS |
| コーディングランキング | 236 | 49.0 | AA |
| 総合ランキング | 328 | 39.0 | AA |
| 科学 | 285 | 46.0 | AA |
ベンチマークスコア (LLM Stats)
(LLM Stats (zeroeval))Agents
Biology
Chemistry
Code
Communication
Creativity
Finance
General
Math
AA評価指数
(Artificial Analysis)LLM Statsカテゴリスコア
(LLM Stats (zeroeval))価格設定
速度
プロバイダー価格ランキング
プロバイダー価格ランキング
11 プロバイダー
このモデルの異なるAPIプロバイダー間の価格を比較。