Qwen3.6 35B A3B (Reasoning)
Description
Qwen3.6-35B-A3B is the first open-weight variant of the Qwen3.6 series, a multimodal Mixture-of-Experts model with 35B total parameters and 3B activated. It pairs a vision encoder with a hybrid 40-layer language model that interleaves Gated DeltaNet linear-attention blocks and Gated Attention blocks (10 × (3 × DeltaNet + 1 × Attention)) over 256 experts (8 routed + 1 shared, expert dim 512). The release prioritizes stability and real-world utility, with substantial gains in agentic coding (frontend workflows, repo-level reasoning) and a new option to preserve reasoning context across turns. Native context length is 262K tokens, extensible to ~1M via YaRN, and the model thinks by default.
Capability Radar
Rankings
| Domain | #Rank | Score | Source |
|---|---|---|---|
| Agentic Capability | 148 | 22.0 | LS |
| Code Ranking | 170 | 60.0 | AA |
| General Ranking | 121 | 67.0 | AA |
| Multimodal Ranking | 19 | 62.0 | LS |
| Science | 152 | 62.0 | AA |
Benchmark Scores (LLM Stats)
(LLM Stats (zeroeval))Agents
Biology
Chemistry
Code
Embodied
Finance
General
Grounding
Healthcare
Long Context
Math
Multimodal
Reasoning
Spatial Reasoning
Vision
AA Evaluation Indices
(Artificial Analysis)LLM Stats Category Scores
(LLM Stats (zeroeval))Pricing
Speed
Provider Price Ranking
Provider Price Ranking
8 providers
Compare pricing across different API providers for this model.