LLMPrice
Updated daily
Z.ai · OpenRouter

Z.ai: GLM 5.3 Prime pricing

GLM-5.3-Prime is the high-speed variant of Z.ai's GLM-5.3, inheriting its full capabilities while delivering 1.5–2× the output throughput through inference acceleration. It supports text input and output with a 1M-token...

OpenRouterOpenRouter API
Input$2.8per 1M tokens
Cached input$0.56per 1M tokens
Output$8.8per 1M tokens
Context1Mmaximum window
Endpoint and source

OpenRouter

Open OpenRouter API source ↗

This is OpenRouter route pricing from the public models API, not a claim about the creator's direct API price.

Tool useReasoningCaching
2,000 input + 500 output tokens$0.01 / request
Calculate a complete workload →