Ling-3.0-flash chat

openrouter
*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...

Capabilities

Context Window 131k tokens
Max Output 16k tokens
Inputs
Outputs

Pricing (per 1M tokens)

Input $0.07
Output $0.22
Cache Read $0.01
Cache Write -

Supported Parameters

frequency_penaltyinclude_reasoninglogit_biasmax_tokensmin_ppresence_penaltyreasoningrepetition_penaltyseedstoptemperaturetool_choicetoolstop_ktop_p