Efficient Qwen model for fast chat, extraction, and high-volume workloads
Benchmarks
| Benchmark | Category | Score | Variant | Provenance | Source | Observed |
|---|---|---|---|---|---|---|
| AA Intelligence Index | intelligence | 6.4 | — | independent | aa | Sep 9, 2026 |
| Median output speed | speed | 100 tok/s | — | independent | aa | Sep 21, 2026 |
| Median time to first token | speed | 2.18s | — | independent | aa | Sep 21, 2026 |
Includes index data from Artificial Analysis.