Faster, lower-cost GLM-5 variant tuned for real-time, high-throughput workloads.
Value rank
#13 of 18
intelligence per dollar in our catalogue
Context rank
#3 tied
198K token window
Independent scores by Artificial Analysis, compared with the strongest models in our catalogue.
Intelligence index
Long context
Where this model earns its keep.
Pricing is live from our platform. Prices per 1M tokens, zero data retention on every request.
| Input price | $1.20 |
| Cache read price | $0.30 |
| Output price | $4.00 |
| Context window | 198K tokens |
| Intelligence / coding index | 26.6 / - |
| Agentic: Terminal-Bench v2.1 / tau2 | - / 99% |
| Long-context reasoning | 72% |
| GPQA / MMLU-Pro | 85% / - |
Close alternatives in the catalogue.
glm-5.1
Intelligence 26.4 · agentic 62%
kimi-k2.7-code
Intelligence 26.3 · agentic 67%
kimi-k2.6
Intelligence 27.5 · agentic 66%
OpenAI-compatible. Switch in one line.
Benchmark data from the Artificial Analysis Intelligence Index, measured independently. Pricing live from the TensorX platform. All inference on EU-sovereign infrastructure with zero data retention.