Z.ai: GLM 5.3 FlashX
Other
Multimodal
Paid
GLM-5.3-FlashX is the high-speed variant of Z.ai's GLM-5.3-Flash, a native multimodal model delivering inference speeds of up to 200 tokens/s. Built on the same hybrid sparse and linear attention architecture...
Parameters
-
Context Window
1,048,576
tokens
Input Price
$0.37
per 1M tokens
Output Price
$1.25
per 1M tokens
Capabilities
Model capabilities and supported modalities
Performance
Reasoning
-
Math
-
Coding
-
Knowledge
-
Modalities
Input Modalities
text,image,video
Output Modalities
text
LLM Price Calculator
Calculate the cost of using this model
$0.000555
$0.003750
Input Cost:$0.000555
Output Cost:$0.003750
Total Cost:$0.004305
Estimated usage: 4,500 tokens
Monthly Cost Estimator
Based on different usage levels
Light Usage
$0.0162
~10 requests
Moderate Usage
$0.1620
~100 requests
Heavy Usage
$1.6200
~1000 requests
Enterprise
$16.2000
~10,000 requests
Note: Estimates based on current token count settings per request.
Last Updated: 2026/09/29
