Qwen3.8-Flash-Next
Descripción
Qwen3.8-Flash-Next is an open-weight experimental preview of the architecture planned for Qwen4, with Hybrid Attention (QSA), Gated Residual, and N-gram Embedding. The card reports 125B parameters with 6B activated, plus 51B n-gram embedding and 4B MTP; Hugging Face BF16 safetensors total 179,999,981,424 parameters (~180B stored). It is a causal LM with a vision encoder for text, image, and video, a native 262,144-token context extensible to 1,000,000 tokens, and thinking on by default (enable_thinking, preserve_thinking, reasoning_effort). This catalog entry is the open-weight checkpoint, not the separate production Qwen3.8-Flash API on Qwen Cloud.
Radar de capacidades
Science se estima a partir de las puntuaciones científicas de LLM Stats o del razonamiento cuando no hay benchmarks científicos dedicados.
Rankings
| Dominio | #Posición | Puntuación | Fuente |
|---|---|---|---|
| Ranking multimodal | 25 | 62.0 | LS |
Puntuaciones de benchmarks (LLM Stats)
(LLM Stats (zeroeval))Agents
Code
Instruction Following
Math
Multimodal
Reasoning
Vision
Índices de evaluación AA
(Artificial Analysis)No hay datos de evaluación AA disponibles
Puntuaciones por categoría LLM Stats
(LLM Stats (zeroeval))Precios
No hay datos de precios disponibles
Velocidad
No hay datos de velocidad disponibles
Ranking de Precios por Proveedor
No hay datos de proveedores disponibles