Llama 3.3 70B Instruct (CoreWeave) API
CoreWeave · serves llama-3.3-70b
Pricing
| Input / 1M | $0.71 |
|---|---|
| Output / 1M | $0.71 |
| Blended / 1M | $1.42 |
| Cached input / 1M | $0.71 |
| Context window | 131,072 tokens |
| Max output | 115,200 tokens |
| License | Open weights · Llama Community License conditional |
| Knowledge cutoff | December 2024 · source |
textfunction-calling
Price verified 2026-09-05 (0 days ago) · sourcecorroborated
Sources
| Source | Input / 1M | Output / 1M | Verified |
|---|---|---|---|
| OpenRouter | $0.71 | $0.71 | 2026-09-05 |
| LiteLLM | $0.71 | $0.71 | 2026-09-03 |
| models.dev | $0.71 | $0.71 | 2026-09-04 |
Two independent sources agree within 20%.
Cheaper providers for llama-3.3-70b
| Provider | Blended / 1M | Endpoint |
|---|---|---|
| DeepInfra | $0.42 | view → |
| Nebius | $0.53 | view → |
| Novita AI | $0.535 | view → |
Compare all 12 providers serving llama-3.3-70b →
Call it
Served via OpenRouter, pinned to CoreWeave — the provider these prices are resolved from. Set OPENROUTER_API_KEY. Provider page.
curl https://openrouter.ai/api/v1/chat/completions \
-H "Authorization: Bearer $OPENROUTER_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "meta-llama/llama-3.3-70b-instruct",
"messages": [{"role": "user", "content": "Hello"}],
"provider": {"only": ["CoreWeave"], "allow_fallbacks": false}
}'from openai import OpenAI # pip install openai
client = OpenAI(
base_url="https://openrouter.ai/api/v1",
api_key="...", # OPENROUTER_API_KEY
)
resp = client.chat.completions.create(
model="meta-llama/llama-3.3-70b-instruct",
messages=[{"role": "user", "content": "Hello"}],
extra_body={"provider": {"only": ["CoreWeave"], "allow_fallbacks": false}},
)
print(resp.choices[0].message.content)Price unchanged since 2026-09-03 — 1 point on record.
Use this data
Raw JSON for this page: llama-3.3-70b.json. Free to reuse under CC BY 4.0 with attribution to apipriceindex.com.