Standard & Paid Price
The locked value is the exact Paid Price after your first top-up; the discount is derived from the two absolute prices
Unlock with a top-up
| Meter | Standard Price | Paid Price | Save |
|---|---|---|---|
| Input price | $1.2000 / 1M tokens | $1.1400 / 1M tokens | -5% |
| Completion price | $4.0000 / 1M tokens | $3.8000 / 1M tokens | -5% |
| Cache read price | $0.2400 / 1M tokens | $0.2400 / 1M tokens | — |
Billing details
Description
Zhipu GLM-5-Turbo — the throughput-oriented GLM-5 variant, optimised for agent workflows: more stable tool and skill invocation across multi-step tasks, better handling of complex long-chain instructions, and scheduled or long-running task execution. 200K context window, up to 128K output tokens, text in and text out. Supports thinking modes, function calling, structured JSON output, streaming, context caching, and MCP tool integration.
Capabilities
- Input modalities
- text
- Output modalities
- text
- Tasks
- chat, reasoning
- Verified features
- streaming
API endpoints
POST /v1/messages(anthropic)POST /v1/chat/completions(openai)