LLMPrice
Updated daily
Z.ai · OpenRouter

Z.ai: GLM 5.3 FlashX pricing

GLM-5.3-FlashX is the high-speed variant of Z.ai's GLM-5.3-Flash, a native multimodal model delivering inference speeds of up to 200 tokens/s. Built on the same hybrid sparse and linear attention architecture...

OpenRouterOpenRouter API
Input$0.37per 1M tokens
Cached input$0.075per 1M tokens
Output$1.25per 1M tokens
Context1.05Mmaximum window
Endpoint and source

OpenRouter

Open OpenRouter API source ↗

This is OpenRouter route pricing from the public models API, not a claim about the creator's direct API price.

Tool useVisionVideo inputReasoningCaching
2,000 input + 500 output tokens$0.001365 / request
Calculate a complete workload →