Back to catalog
Publisher : Z.aiFamily : GLMAvailableNo subscriptionHigh availability

GLM 5.3 Flash API

Model ID : glm-5.3-flash

Model ID

glm-5.3-flash

OpenAI-compatible API

Context window

1,048,576 tokens

Maximum output

131,072 tokens

Input / 1M tokens

0.15 USD

Output / 1M tokens

0.50 USD

Cache read / 1M tokens

$0.03

Cache write / 1M tokens

$0.15

Cache reads reuse stored tokens; cache writes create the provider cache.

Access to this model

Authorized plans : bronze, gold, silver, pro, business

This model is also available with credits without a subscription.

Published capabilities

  • Vision
  • Tool calling
  • Streaming
  • Structured
  • Reasoning

Minimal request

curl https://api.routerlab.ch/v1/chat/completions \
  -H "Authorization: Bearer $ROUTERLAB_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"glm-5.3-flash","max_tokens":128,"messages":[{"role":"user","content":"Hello RouterLab"}]}'