LLMPrice
Updated daily
Z.ai · OpenRouter

Z.ai: GLM 5.3 Flash (batch) pricing

GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...

OpenRouterOpenRouter API
Input$0.15per 1M tokens
Cached input$0.03per 1M tokens
Output$0.5per 1M tokens
Context1.05Mmaximum window
Endpoint and source

OpenRouter

Open OpenRouter API source ↗

This is OpenRouter route pricing from the public models API, not a claim about the creator's direct API price.

Tool useVisionVideo inputReasoningCachingHugging Face
2,000 input + 500 output tokens$0.00055 / request
Calculate a complete workload →