LLMPrice
Updated daily
Inference Net · OpenRouter

Inference.net: Schematron V2 Turbo pricing

Schematron V2 Turbo is a 3B-parameter HTML-to-JSON extraction model from Inference.net. It prioritizes throughput for high-volume extraction workloads. Extraction instructions must be supplied through a JSON schema in response_format rather...

OpenRouterOpenRouter API
Input$0.03per 1M tokens
Cached input$0.03per 1M tokens
Output$0.15per 1M tokens
Context128Kmaximum window
Endpoint and source

OpenRouter

Open OpenRouter API source ↗

This is OpenRouter route pricing from the public models API, not a claim about the creator's direct API price.

CachingHugging Face
2,000 input + 500 output tokens$0.000135 / request
Calculate a complete workload →