MODEL SELECTION TOOLS
Find the right fit.
Compare capabilities and rates, then estimate costs for your own workload.
Your estimated workload
Calculated in each model’s billing unit. Models with different capabilities are not interchangeable.
| Model & cost | GLM 5.2 FP8 Primary LLM for chat, reasoning, and coding workloads. |
|---|---|
| Estimated costUSD | $1.68Based on input & output tokens |
| Input / unit rate | $0.93per 1M input tokens |
| Output rate | $3.00per 1M output tokens |
| Context window | 131K |
| Maximum output | 16K |
| Streaming | ✓ Supported |
| Tool calling | ✓ Supported |
| Structured output | ✓ Supported |
| Get started | Configure integration ↗ |
Based on published catalog rates and specifications, not performance benchmarks or a live bill. See service status for model availability. Check status ↗