z-ai/glm-5
GLM-5 is Z.ai’s flagship open-source foundation model engineered for complex systems design and long-horizon agent workflows. Built for expert developers, it delivers production-grade performance on large-scale programming tasks, rivaling leading...
| Provider | In $/M | Out $/M | Context | Uptime |
|---|---|---|---|---|
| StreamLakefp8 | $0.600 | $1.92 | 198K | 99.8% |
| GMICloudfp8 | $0.600 | $1.92 | 203K | 99.6% |
| DeepInfrafp4 | $0.600 | $2.08 | 203K | 100% |
| Baidufp8cache | $0.700 | $2.24 | 203K | 98.9% |
| DigitalOcean | $0.750 | $2.40 | 64K | 97.6% |
| SiliconFlowfp8 | $0.950 | $2.55 | 205K | 99.7% |
| AtlasCloudfp8 | $0.950 | $3.15 | 203K | 100% |
| Amazon Bedrock | $1.00 | $3.20 | 203K | 100% |
| Novitafp8 | $1.00 | $3.20 | 203K | 100% |
| Z.AIfp8 | $1.00 | $3.20 | 203K | 99.2% |
| Parasailfp8 | $1.00 | $3.20 | 203K | 98.2% |
| Venicefp8 | $1.00 | $3.20 | 198K | — |
| Phala | $1.20 | $3.50 | 203K | — |
GLM 5 costs $0.950 per million input tokens and $2.55 per million output tokens via OpenRouter, making it 218th cheapest of 332 paid models.
GLM 5 scores 39.5 on the Artificial Analysis Intelligence Index, ranking 47th of 182 benchmarked models, with a GPQA Diamond score of 82%.
GLM 5 supports a 205K-token context window and can output up to 131K tokens. It accepts text input.