Z.AI / zai/glm-5.3-flash

GLM-5.3-Flash - access through LLMTR

GLM-5.3-Flash processes images such as screenshots, interfaces, and diagrams alongside code and text, with function calling for agent workflows. Its 1M context window leaves room for long working sessions. Reasoning runs on every response; plain requests use low, while harder tasks can select high or max.

Technical specifications

Canonical IDzai/glm-5.3-flash
ProviderZ.AI
Context window1,000,000 tokens
OperationsCHAT_COMPLETIONS
Modalitiestext, image

Pricing

An 8% platform margin applies to credit top-ups; model usage prices are not separately marked up.

OperationMetricUnitPrice
CHAT_COMPLETIONSINPUT_TEXTPER_1M_TOKENS$0.075000
CHAT_COMPLETIONSCACHE_READPER_1M_TOKENS$0.015000
CHAT_COMPLETIONSOUTPUT_TEXTPER_1M_TOKENS$0.250000

Example usage

With existing OpenAI SDK flows, change only the base URL and model identifier.

curl https://llmtr.com/v1/chat/completions \
  -H "Authorization: Bearer llmtr-your_key" \
  -H "Content-Type: application/json" \
  -d '{"model":"zai/glm-5.3-flash","messages":[{"role":"user","content":"Hello"}]}'

Related models