Qwen3.7 Plus
Description
Qwen3.7-Plus is Alibaba Cloud Qwen Team's multimodal agent model that unifies vision and language into a single agent foundation. Built on the Qwen3.7 text backbone, it operates as a multimodal interactive hybrid agent—perceiving real-world scenes, reading screens and operating GUIs, writing code from visual references, navigating mobile apps end-to-end, and answering search-augmented visual questions—while blending GUI and CLI interactions within a single agent loop. It is a versatile coding agent and productivity assistant with full-modality input, generalizing across scaffolds such as Claude Code, OpenClaw, and Qwen Code. Features a 1 million token context window, up to 65,536 output tokens, always-on thinking, and a preserve_thinking mode for agentic tasks. Available via Alibaba Cloud Model Studio (DashScope).
Capability Radar
Rankings
| Domain | #Rank | Score | Source |
|---|---|---|---|
| Agentic Capability | 30 | 53.0 | LS |
| Code Ranking | 99 | 74.0 | AA |
| General Ranking | 67 | 78.0 | AA |
| Multimodal Ranking | 2 | 75.0 | LS |
| Science | 71 | 78.0 | AA |
Benchmark Scores (LLM Stats)
(LLM Stats (zeroeval))Agents
Chat
Code
General
Instruction Following
Language
Long Context
Math
Multimodal
Reasoning
Tool Calling
Video
Vision
AA Evaluation Indices
(Artificial Analysis)LLM Stats Category Scores
(LLM Stats (zeroeval))Pricing
Speed
Provider Price Ranking
Provider Price Ranking
19 providers
Compare pricing across different API providers for this model.