Provider
Z AiModel detail
z-ai/glm-5.3
Opt-in LINE comparison pilot; fallback follows the Router's configured policy. Catalog prices are list prices, not a promise of free credits. Qwen cache rate is for implicit cache; explicit-cache creation/read is not enabled by this route.
Context
128Ktokens
Max output
33Ktokens
Default endpoint
fireworks/line-evaluationInput price
$1.4000per 1M tokens
Output price
$4.4000per 1M tokens
Energy estimate
75 Whper 1M tokens
Fallback models
1OpenRouter-compatible metadata
- Model ID
z-ai/glm-5.3- Canonical slug
z-ai/glm-5.3- Owned by
- z-ai
- API
- zevrouter
- Input modalities
- text
- Output modalities
- text
- Tokenizer
- Other
- Instruct type
- chat
Pricing JSON
{
"prompt": "0.0000014",
"completion": "0.0000044",
"request": "0",
"image": "0",
"audio": "0",
"web_search": "0",
"internal_reasoning": "0"
}
Parameters
Supported parameters
temperaturetop_pmax_tokensmax_completion_tokenstoolstool_choicestreamstream_optionsreasoning_effort
Routing
Provider endpoints
| Endpoint | Region | Status | Latency | Carbon intensity | CFE |
|---|---|---|---|---|---|
fireworks/line-evaluationfireworks evaluation |
provider-managed | default | -- | -- | -- |
Fallback chain
{
"model": "z-ai/glm-5.3",
"models": ["z-ai/glm-5.3", "google/gemini-3.7-flash"],
"messages": [{ "role": "user", "content": "Hello" }]
}
API references
curl "https://router.zev.city/api/v1/models"
curl "https://router.zev.city/api/v1/models/z-ai/glm-5.3/endpoints"