inkling-small API — compare providers

The same model (inkling-small) is served by 3 providers; DeepInfra is the cheapest at $1.65/1M blended — the priciest (Together AI) costs 3% more.

RankProviderInput /1MOutput /1MBlended /1M
1DeepInfra$0.45$1.2$1.65 ← cheapest
2BaseTen$0.5$1.2$1.7
3Together AI$0.5$1.2$1.7

Prices are per 1M tokens (USD). "Blended" = input + output, for coarse ranking. Same underlying model, 3 serving providers, 3% spread top to bottom. Context window served: 1,048,576 tokens (262,144 max output) on all 3 priced endpoints — see how it ranks in biggest context windows. See how this compares across the catalogue in same model, different price.

inkling-small spec sheetmaker-declared

What the maker declares about the model itself — read from its own weights repository and documentation, not from any host. What each endpoint actually serves is in the pricing above: the two can differ, and both are true.

Context window
to verify
Parameters
to verify
Active experts
to verify
Weight precision
to verify
License
to verifyScope: the licence on the published weights, and nothing else. It is not the maker's acceptable-use policy, which is a separate document, and it is not the contract you sign with whichever host you call — hosts set their own terms, which this index does not read.
Knowledge cutoff
to verify
Modalities
to verify
Max output (official)
to verify
Release date
to verify≈ July 30, 2026third-party catalogue — not the maker's figure
Maker lifecycle status
to verify

no maker source read yet · 0 of 10 fields published by the maker

Cost per 1,000 requests by workload

What each provider actually bills for a representative job, not just the sticker price.

ProviderChatbot
1000 in / 500 out
RAG / long context
8000 in / 500 out
Batch summarize
4000 in / 1000 out
DeepInfra$1.05$4.20$3.00
BaseTen$1.10$4.60$3.20
Together AI$1.10$4.60$3.20

Estimate your own workload

Input tokens/request:   Output tokens/request:   Requests:

ProviderEstimated cost (USD)