Ling 3.0 Flash
inclusionAI · inclusionai/ling-3.0-flash
*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...
Ling 3.0 Flash ranks #51 of 348 models in the default AI Benchmark Hub order (quality first, then cost). At $0.04 blended per million tokens, it is cheaper than 97% of models with published API pricing. Its 262,144-token context window is about the catalog median (262,144 tokens).
Context
Max output: 32768
Pricing
Output / 1M: 0.06
Blend / 1M: 0.04
Quality
Provider
Moderated: no
Monthly cost at this blended rate
Estimated API spend if every token is billed at $0.04per million (the blend of this model's input and output prices). Real bills depend on the input/output mix.
| Tokens / month | Est. cost (USD) |
|---|---|
| 1,000,000 | $0.04 |
| 10,000,000 | $0.42 |
| 100,000,000 | $4.20 |
Related models
Same lab, similar price, or similar quality — useful next comparisons.
More from inclusionAI
- Ling 3.0 Flash Sante$0.00 / 1M
- Ling 3.0 Flash Fin$0.12 / 1M
Closest API price
- Llama 3 8B Lunaris$0.04 / 1M
- Mistral Nemo$0.02 / 1M
- MythoMax 13B$0.06 / 1M
- Nex-N2-Mini$0.06 / 1M
- Granite 4.0 Micro$0.06 / 1M