Back to all models
inclusionai

Ling-3.0-flash

*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...

inclusionai/ling-3.0-flash

Context Size

262.144K

Input Price

3,150 pts/M

Output Price

9,450 pts/M


Architecture

Text

Supported Parameters

frequency_penaltyinclude_reasoninglogit_biaslogprobsmax_tokensmin_ppresence_penaltyreasoningrepetition_penaltyresponse_formatseedstoptemperaturetool_choicetoolstop_ktop_logprobstop_p

Details

TokenizerOther
Max Completion32,768 tokens
Provider Context262.144K tokens
ModeratedNo
CreatedJul 23, 2026