Ling 3.0 Flash

inclusionAI · inclusionai/ling-3.0-flash

← Back to leaderboard

*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...

Ling 3.0 Flash ranks #51 of 348 models in the default AI Benchmark Hub order (quality first, then cost). At $0.04 blended per million tokens, it is cheaper than 97% of models with published API pricing. Its 262,144-token context window is about the catalog median (262,144 tokens).

open weightstexttext->text

Context

Max context: 262144
Max output: 32768

Pricing

Input / 1M: 0.02
Output / 1M: 0.06
Blend / 1M: 0.04

Quality

Quality index:

Provider

Provider: inclusionAI
Moderated: no

Monthly cost at this blended rate

Estimated API spend if every token is billed at $0.04per million (the blend of this model's input and output prices). Real bills depend on the input/output mix.

Tokens / monthEst. cost (USD)
1,000,000$0.04
10,000,000$0.42
100,000,000$4.20

Related models

Same lab, similar price, or similar quality — useful next comparisons.

More from inclusionAI

Closest API price