Inkling Small
Thinkingmachines · thinkingmachines/inkling-small
Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B total. It is positioned as the smaller, more efficient member of...
Inkling Small ranks #184 of 348 models in the default AI Benchmark Hub order (quality first, then cost). At $0.82 blended per million tokens, it is cheaper than 56% of models with published API pricing. Its 524,288-token context window is 2.0× the catalog median (262,144 tokens).
Context
Max output: 262144
Pricing
Output / 1M: 1.20
Blend / 1M: 0.82
Quality
Provider
Moderated: no
Monthly cost at this blended rate
Estimated API spend if every token is billed at $0.82per million (the blend of this model's input and output prices). Real bills depend on the input/output mix.
| Tokens / month | Est. cost (USD) |
|---|---|
| 1,000,000 | $0.82 |
| 10,000,000 | $8.25 |
| 100,000,000 | $82.50 |
Related models
Same lab, similar price, or similar quality — useful next comparisons.
More from Thinkingmachines
- Inkling$2.52 / 1M
Closest API price
- Perceptron Mk1$0.82 / 1M
- Qwen2.5 Coder 32B Instruct$0.83 / 1M
- ERNIE 4.5 VL 424B A47B$0.83 / 1M
- Qwen3.7 Plus$0.80 / 1M · Q 78.7
- R1 Distill Llama 70B$0.80 / 1M