Qwen vision-language model for visual reasoning, documents, and agent tasks
Benchmarks
| Benchmark | Category | Score | Variant | Provenance | Source | Observed |
|---|---|---|---|---|---|---|
| AA Agentic Index | agentic | 16.2 | non-reasoning | independent | aa | Aug 7, 2026 |
| 7.6 | — | mirror | openrouter | Sep 20, 2026 | ||
| 7.6 | — | independent | aa | Sep 20, 2026 | ||
| AA Coding Index | coding | 45.7 | — | mirror | openrouter | Jul 23, 2026 |
| 43.3 | non-reasoning | independent | aa | Jul 24, 2026 | ||
| 45.7 | — | independent | aa | Jul 24, 2026 | ||
| AA Intelligence Index | intelligence | 17.7 | non-reasoning | independent | aa | Sep 9, 2026 |
| 15.6 | — | mirror | openrouter | Sep 20, 2026 | ||
| 15.6 | — | independent | aa | Sep 20, 2026 | ||
| LMArena Text | preference | 1418 | — | crowd | lmarena | Sep 13, 2026 |
| LMArena Vision | preference | 1245 | — | crowd | lmarena | Sep 13, 2026 |
| LMArena WebDev | preference | 1358 | — | crowd | lmarena | Sep 11, 2026 |
| Median output speed | speed | 145 tok/s | non-reasoning | independent | aa | Sep 21, 2026 |
| 128 tok/s | — | independent | aa | Sep 21, 2026 | ||
| Median time to first token | speed | 2.34s | non-reasoning | independent | aa | Sep 21, 2026 |
| 2.35s | — | independent | aa | Sep 21, 2026 |
Includes index data from Artificial Analysis.
Providers
Live operational stats per hosting endpoint, via OpenRouter.
| Provider | Quant | Context | $/1M in · out | Uptime 24h | Latency | Throughput |
|---|---|---|---|---|---|---|
| Alibaba | — | 262K | $0.26 · $2.08 | 98.38% | — | — |
| AtlasCloud | fp8 | 262K | $0.3 · $2.4 | 98.71% | — | — |
| DeepInfra | fp4 | 262K | $0.29 · $2.4 | 99.56% | — | — |
| Novita | bf16 | 262K | $0.4 · $3.2 | 99.81% | — | — |
| SiliconFlow | fp8 | 262K | $0.26 · $2.08 | 87.74% | — | — |