Llama 3.1 8B Instruct (CoreWeave) API
CoreWeave · serves llama-3.1-8b
Pricing
| Input / 1M | $0.22 |
|---|---|
| Output / 1M | $0.22 |
| Blended / 1M | $0.44 |
| Cached input / 1M | $0.22 |
| Context window | 131,072 tokens |
| Max output | 117,964 tokens |
| License | Open weights · Llama Community License conditional |
| Knowledge cutoff | December 2024 · source |
textfunction-calling
Price verified 2026-09-05 (0 days ago) · sourcecorroborated
Sources
| Source | Input / 1M | Output / 1M | Verified |
|---|---|---|---|
| OpenRouter | $0.22 | $0.22 | 2026-09-05 |
| LiteLLM | $0.22 | $0.22 | 2026-09-03 |
| models.dev | $0.22 | $0.22 | 2026-09-04 |
Two independent sources agree within 20%.
Cheaper providers for llama-3.1-8b
| Provider | Blended / 1M | Endpoint |
|---|---|---|
| DeepInfra | $0.06 | view → |
| Novita AI | $0.07 | view → |
| Groq | $0.13 | view → |
Compare all 5 providers serving llama-3.1-8b →
Call it
Served via OpenRouter, pinned to CoreWeave — the provider these prices are resolved from. Set OPENROUTER_API_KEY. Provider page.
curl https://openrouter.ai/api/v1/chat/completions \
-H "Authorization: Bearer $OPENROUTER_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "meta-llama/llama-3.1-8b-instruct",
"messages": [{"role": "user", "content": "Hello"}],
"provider": {"only": ["CoreWeave"], "allow_fallbacks": false}
}'from openai import OpenAI # pip install openai
client = OpenAI(
base_url="https://openrouter.ai/api/v1",
api_key="...", # OPENROUTER_API_KEY
)
resp = client.chat.completions.create(
model="meta-llama/llama-3.1-8b-instruct",
messages=[{"role": "user", "content": "Hello"}],
extra_body={"provider": {"only": ["CoreWeave"], "allow_fallbacks": false}},
)
print(resp.choices[0].message.content)Price unchanged since 2026-09-03 — 1 point on record.
Use this data
Raw JSON for this page: llama-3.1-8b.json. Free to reuse under CC BY 4.0 with attribution to apipriceindex.com.