LLMPrice
Updated daily
Inclusion AI · Model identity

Ling 3.0 Flash

*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...

1pricing route
Context262Klargest published route limit
Max output33Klargest published route limit
InputsTextfrom route metadata
OutputsNot publishedfrom route metadata
Capabilities

What published route metadata supports

CachingHugging FaceReasoningStructured outputTool use
Input modalities
Text
Supported API parameters
frequency_penaltyinclude_reasoninglogit_biaslogprobsmax_tokensmin_ppresence_penaltyreasoningrepetition_penaltyresponse_formatseedstoptemperaturetool_choicetoolstop_ktop_logprobstop_p
Model facts

Only source-backed fields are filled

Release dateNot independently verified
Knowledge cutoffNot independently verified
Open weightsNot independently verified
OpenRouter catalog date2026-07-23

LLMPrice does not infer a release date, knowledge cutoff, or open-weight status from a model name, Hugging Face link, or pricing route.

Could this workload cost less?Compare Ling 3.0 Flash with cheaper models using the same workload.
Find cheaper alternatives →
Pricing routes

1 ways this model appears in the catalog

Calculate this model →
Endpoint / routeModel IDInputCachedOutputContextSource
OpenRouterStandardinclusionai/ling-3.0-flash$0.021$0.0042$0.063262KOpenRouter API ↗
Published repository IDs

Hugging Face references

Presence of a Hugging Face ID is not treated as proof that model weights are open.

Price history

Tracking since Aug 26, 2026

No verified price changes have been recorded since tracking began on Aug 26, 2026. The Aug 26, 2026 catalog is the baseline; LLMPrice does not invent or backfill older prices.

See all verified price changes →