inkling-small API — compare providers
The same model (inkling-small) is served by 3 providers; DeepInfra is the cheapest at $1.65/1M blended — the priciest (Together AI) costs 3% more.
| Rank | Provider | Input /1M | Output /1M | Blended /1M |
|---|---|---|---|---|
| 1 | DeepInfra | $0.45 | $1.2 | $1.65 ← cheapest |
| 2 | BaseTen | $0.5 | $1.2 | $1.7 |
| 3 | Together AI | $0.5 | $1.2 | $1.7 |
Prices are per 1M tokens (USD). "Blended" = input + output, for coarse ranking. Same underlying model, 3 serving providers, 3% spread top to bottom. Context window served: 1,048,576 tokens (262,144 max output) on all 3 priced endpoints — see how it ranks in biggest context windows. See how this compares across the catalogue in same model, different price.
inkling-small spec sheetmaker-declared
What the maker declares about the model itself — read from its own weights repository and documentation, not from any host. What each endpoint actually serves is in the pricing above: the two can differ, and both are true.
- Context window
- to verify
- Parameters
- to verify
- Active experts
- to verify
- Weight precision
- to verify
- License
- to verifyScope: the licence on the published weights, and nothing else. It is not the maker's acceptable-use policy, which is a separate document, and it is not the contract you sign with whichever host you call — hosts set their own terms, which this index does not read.
- Knowledge cutoff
- to verify
- Modalities
- to verify
- Max output (official)
- to verify
- Release date
- to verify≈ July 30, 2026third-party catalogue — not the maker's figure
- Maker lifecycle status
- to verify
no maker source read yet · 0 of 10 fields published by the maker
Cost per 1,000 requests by workload
What each provider actually bills for a representative job, not just the sticker price.
| Provider | Chatbot 1000 in / 500 out | RAG / long context 8000 in / 500 out | Batch summarize 4000 in / 1000 out |
|---|---|---|---|
| DeepInfra | $1.05 | $4.20 | $3.00 |
| BaseTen | $1.10 | $4.60 | $3.20 |
| Together AI | $1.10 | $4.60 | $3.20 |
Estimate your own workload
Input tokens/request: Output tokens/request: Requests:
| Provider | Estimated cost (USD) |
|---|