Skip to content
Z.aiNew
z-ai/glm-5.3-flashx

About this model

GLM-5.3-FlashX is the high-speed variant of Z.ai's GLM-5.3-Flash, a native multimodal model delivering inference speeds of up to 200 tokens/s. Built on the same hybrid sparse and linear attention architecture...

Make a request

OpenAI-compatible: set the base URL and your Promptix API key, keep your existing SDK.

curl https://promptix.tn/api/v1/chat/completions \  -H "Authorization: Bearer $PROMPTIX_API_KEY" \  -H "Content-Type: application/json" \  -d '{    "model": "z-ai/glm-5.3-flashx",    "messages": [      {"role": "user", "content": "Hello! What can you do?"}    ]  }'

Pricing

In TND per 1M tokens. VAT and the platform fee are added at top-up.

Input tokens

1.432 TND

per 1M tokens

Output tokens

4.839 TND

per 1M tokens

1,000 requests with 1K input and 500 output tokens cost about 3.852 TND.

Specifications

Provider
Z.ai
Context window
1,048,576 tokens
Input modality
Text, Image
Added
Sep 18, 2026
Supported parameters
include_reasoningmax_tokensreasoningreasoning_effortresponse_formattemperaturetool_choicetools