Compact GPT model for low-latency assistance and high-volume workloads
Benchmarks
| Benchmark | Category | Score | Variant | Provenance | Source | Observed |
|---|---|---|---|---|---|---|
| AA Coding Index | coding | 21.5 | — | mirror | openrouter | Jul 23, 2026 |
| LMArena Text | preference | 1272 | — | crowd | lmarena | Sep 13, 2026 |
| LMArena Vision | preference | 1090 | — | crowd | lmarena | Sep 13, 2026 |
Includes index data from Artificial Analysis.
Providers
Live operational stats per hosting endpoint, via OpenRouter.
| Provider | Quant | Context | $/1M in · out | Uptime 24h | Latency | Throughput |
|---|---|---|---|---|---|---|
| OpenAI | — | 128K | $10 · $30 | 99.87% | — | — |