Skip to content
Z.ai
z-ai/glm-5.3-flash

About this model

GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...

Make a request

OpenAI-compatible: set the base URL and your Promptix API key, keep your existing SDK.

curl https://promptix.tn/api/v1/chat/completions \  -H "Authorization: Bearer $PROMPTIX_API_KEY" \  -H "Content-Type: application/json" \  -d '{    "model": "z-ai/glm-5.3-flash",    "messages": [      {"role": "user", "content": "Hello! What can you do?"}    ]  }'

Pricing

In TND per 1M tokens. VAT and the platform fee are added at top-up.

Input tokens

0.155 TND

per 1M tokens

Output tokens

1.935 TND

per 1M tokens

1,000 requests with 1K input and 500 output tokens cost about 1.123 TND.

Specifications

Provider
Z.ai
Context window
1,310,720 tokens
Input modality
Text, Image
Added
Aug 26, 2026
Supported parameters
frequency_penaltyinclude_reasoninglogit_biaslogprobsmax_tokensmin_pparallel_tool_callspresence_penalty