Catalog capabilities
- Vision input
- Tool calling
- Streaming
- Structured output
- Reasoning
RouterLab publishes this GLM model under the model ID glm-5.3-flash-trial. The facts below come from the active RouterLab catalog.
Model ID
glm-5.3-flash-trial
Context window
1,048,576 tokens
Maximum output
131,072 tokens
Input / 1M tokens
$0.01
Output / 1M tokens
$0.01
Cache read / 1M tokens
$0.001
Cache write / 1M tokens
$0.001
Cache read reuses tokens already stored in the provider cache. Cache write creates or stores new cache tokens.
curl https://api.routerlab.ch/v1/chat/completions \
-H "Authorization: Bearer $ROUTERLAB_API_KEY" \
-H "Content-Type: application/json" \
-d '{"model":"glm-5.3-flash-trial","messages":[{"role":"user","content":"Hello RouterLab"}]}'Related models