Cost-efficient GPT-5.6 model for fast, high-volume workloads
Benchmarks
| Benchmark | Category | Score | Variant | Provenance | Source | Observed |
|---|---|---|---|---|---|---|
| AA Agentic Index | agentic | 45.6 | — | mirror | openrouter | Jul 23, 2026 |
| 45.6 | — | independent | aa | Jul 24, 2026 | ||
| AA Coding Index | coding | 71.4 | — | mirror | openrouter | Jul 23, 2026 |
| 71.4 | — | independent | aa | Jul 24, 2026 | ||
| AA Intelligence Index | intelligence | 51.2 | — | mirror | openrouter | Jul 23, 2026 |
| 51.2 | — | independent | aa | Jul 24, 2026 | ||
| Median output speed | speed | 173 tok/s | — | independent | aa | Jul 24, 2026 |
| Median time to first token | speed | 126.14s | — | independent | aa | Jul 24, 2026 |
Includes index data from Artificial Analysis.
Providers
Live operational stats per hosting endpoint, via OpenRouter.
| Provider | Quant | Context | $/1M in · out | Uptime 24h | Latency | Throughput |
|---|---|---|---|---|---|---|
| Azure | — | 1.1M | $1 · $6 | 99.99% | — | — |
| Azure | — | 1.1M | $1.1 · $6.6 | 100.00% | — | — |
| OpenAI | — | 1.1M | $1 · $6 | 99.47% | — | — |
| OpenAI | — | 1.1M | $0.5 · $3 | 99.47% | — | — |
| OpenAI | — | 1.1M | $2 · $12 | 99.47% | — | — |