Skip to content

Z.ai: GLM 5.3 Flash (batch)

Z.ai
z-ai/glm-5.3-flash:batch

About this model

GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...

Make a request

OpenAI-compatible: set the base URL and your Promptix API key, keep your existing SDK.

curl https://promptix.tn/api/v1/chat/completions \  -H "Authorization: Bearer $PROMPTIX_API_KEY" \  -H "Content-Type: application/json" \  -d '{    "model": "z-ai/glm-5.3-flash:batch",    "messages": [      {"role": "user", "content": "Hello! What can you do?"}    ]  }'

Pricing

In TND per 1M tokens. VAT and the platform fee are added at top-up.

Input tokens

0.232 TND

per 1M tokens

Output tokens

0.774 TND

per 1M tokens

1,000 requests with 1K input and 500 output tokens cost about 0.619 TND.

Specifications

Provider
Z.ai
Context window
1,048,576 tokens
Input modality
Text, Image
Added
Aug 26, 2026
Supported parameters
frequency_penaltyinclude_reasoninglogit_biasmax_tokensmin_ppresence_penaltyreasoningreasoning_effort