Z.ai: GLM 5.3 FlashX
Z.aiNew
z-ai/glm-5.3-flashxAbout this model
GLM-5.3-FlashX is the high-speed variant of Z.ai's GLM-5.3-Flash, a native multimodal model delivering inference speeds of up to 200 tokens/s. Built on the same hybrid sparse and linear attention architecture...
Make a request
OpenAI-compatible: set the base URL and your Promptix API key, keep your existing SDK.
curl https://promptix.tn/api/v1/chat/completions \ -H "Authorization: Bearer $PROMPTIX_API_KEY" \ -H "Content-Type: application/json" \ -d '{ "model": "z-ai/glm-5.3-flashx", "messages": [ {"role": "user", "content": "Hello! What can you do?"} ] }'Pricing
In TND per 1M tokens. VAT and the platform fee are added at top-up.
Input tokens
1.432 TNDper 1M tokens
Output tokens
4.839 TNDper 1M tokens
1,000 requests with 1K input and 500 output tokens cost about 3.852 TND.
Specifications
- Provider
- Z.ai
- Context window
- 1,048,576 tokens
- Input modality
- Text, Image
- Added
- Sep 18, 2026
- Supported parameters
include_reasoningmax_tokensreasoningreasoning_effortresponse_formattemperaturetool_choicetools