Z.AI / zai/glm-5.3-flash
GLM-5.3-Flash - access through LLMTR
GLM-5.3-Flash processes images such as screenshots, interfaces, and diagrams alongside code and text, with function calling for agent workflows. Its 1M context window leaves room for long working sessions. Reasoning runs on every response; plain requests use low, while harder tasks can select high or max.
Technical specifications
| Canonical ID | zai/glm-5.3-flash |
|---|---|
| Provider | Z.AI |
| Context window | 1,000,000 tokens |
| Operations | CHAT_COMPLETIONS |
| Modalities | text, image |
Pricing
An 8% platform margin applies to credit top-ups; model usage prices are not separately marked up.
| Operation | Metric | Unit | Price |
|---|---|---|---|
| CHAT_COMPLETIONS | INPUT_TEXT | PER_1M_TOKENS | $0.075000 |
| CHAT_COMPLETIONS | CACHE_READ | PER_1M_TOKENS | $0.015000 |
| CHAT_COMPLETIONS | OUTPUT_TEXT | PER_1M_TOKENS | $0.250000 |
Example usage
With existing OpenAI SDK flows, change only the base URL and model identifier.
curl https://llmtr.com/v1/chat/completions \
-H "Authorization: Bearer llmtr-your_key" \
-H "Content-Type: application/json" \
-d '{"model":"zai/glm-5.3-flash","messages":[{"role":"user","content":"Hello"}]}'
Related models
- GLM-5.3 zai/glm-5.3
- GLM-5.2 zai/glm-5.2
- GLM-5.1 zai/glm-5.1
- GLM-5 zai/glm-5
- GLM-5-Turbo zai/glm-5-turbo
- GLM-5V-Turbo zai/glm-5v-turbo
- GLM-4.7 zai/glm-4.7
- GLM-4.7-FlashX zai/glm-4.7-flashx