Model detail

z-ai/glm-5.3

Opt-in LINE comparison pilot; fallback follows the Router's configured policy. Catalog prices are list prices, not a promise of free credits. Qwen cache rate is for implicit cache; explicit-cache creation/read is not enabled by this route.

Provider
Z Ai
Context
128K
tokens
Max output
33K
tokens
Default endpoint
fireworks/line-evaluation
Input price
$1.4000
per 1M tokens
Output price
$4.4000
per 1M tokens
Energy estimate
75 Wh
per 1M tokens
Fallback models
1

OpenRouter-compatible metadata

Model ID
z-ai/glm-5.3
Canonical slug
z-ai/glm-5.3
Owned by
z-ai
API
zevrouter
Input modalities
text
Output modalities
text
Tokenizer
Other
Instruct type
chat

Pricing JSON

Parameters

Supported parameters

9 parameters
temperaturetop_pmax_tokensmax_completion_tokenstoolstool_choicestreamstream_optionsreasoning_effort

Routing

Provider endpoints

Inspect Endpoints
EndpointRegionStatusLatencyCarbon intensityCFE
fireworks/line-evaluation
fireworks evaluation
provider-managed default -- -- --

Fallback chain

google/gemini-3.7-flash

{
  "model": "z-ai/glm-5.3",
  "models": ["z-ai/glm-5.3", "google/gemini-3.7-flash"],
  "messages": [{ "role": "user", "content": "Hello" }]
}

API references

curl "https://router.zev.city/api/v1/models"
curl "https://router.zev.city/api/v1/models/z-ai/glm-5.3/endpoints"