Compact GPT model for low-latency assistance and high-volume workloads
Benchmarks
| Benchmark | Category | Score | Variant | Provenance | Source | Observed |
|---|---|---|---|---|---|---|
| AA Coding Index | coding | 21.5 | — | mirror | openrouter | Jul 23, 2026 |
| LMArena Vision | preference | 1090 | — | crowd | lmarena | Jul 21, 2026 |
Includes index data from Artificial Analysis.
Providers
Live operational stats per hosting endpoint, via OpenRouter.
| Provider | Quant | Context | $/1M in · out | Uptime 24h | Latency | Throughput |
|---|---|---|---|---|---|---|
| OpenAI | — | 128K | $10 · $30 | 100.00% | — | — |