Qwen3.8 2.4T A95B (Venice) API
Venice · serves qwen3.8-2.4t-a95b
Pricing
| Input / 1M | $2.50 |
|---|---|
| Output / 1M | $7.50 |
| Blended / 1M | $10 |
| Cached input / 1M | $0.31 |
| Context window | 1,048,576 tokens |
| Max output | 65,536 tokens |
| License | Open weights · Apache 2.0 permissive |
| Knowledge cutoff | September 2025 · source |
textfunction-calling
Price verified 2026-09-05 (0 days ago) · sourcecorroborated
Sources
| Source | Input / 1M | Output / 1M | Verified |
|---|---|---|---|
| OpenRouter | $2.50 | $7.50 | 2026-09-05 |
| LiteLLM | $2.00 | $6.00 | 2026-09-03 |
| models.dev | $2.50 | $6.25 | 2026-09-04 |
Two independent sources agree within 20%.
Cheaper providers for qwen3.8-2.4t-a95b
| Provider | Blended / 1M | Endpoint |
|---|---|---|
| Novita AI | $8 | view → |
| Alibaba | $8 | view → |
| SiliconFlow | $8 | view → |
Compare all 7 providers serving qwen3.8-2.4t-a95b →
Call it
Served via OpenRouter, pinned to Venice — the provider these prices are resolved from. Set OPENROUTER_API_KEY. Provider page.
curl https://openrouter.ai/api/v1/chat/completions \
-H "Authorization: Bearer $OPENROUTER_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "qwen/qwen3.8-2.4t-a95b",
"messages": [{"role": "user", "content": "Hello"}],
"provider": {"only": ["Venice"], "allow_fallbacks": false}
}'from openai import OpenAI # pip install openai
client = OpenAI(
base_url="https://openrouter.ai/api/v1",
api_key="...", # OPENROUTER_API_KEY
)
resp = client.chat.completions.create(
model="qwen/qwen3.8-2.4t-a95b",
messages=[{"role": "user", "content": "Hello"}],
extra_body={"provider": {"only": ["Venice"], "allow_fallbacks": false}},
)
print(resp.choices[0].message.content)Price unchanged since 2026-09-03 — 1 point on record.
Use this data
Raw JSON for this page: qwen3.8-2.4t-a95b.json. Free to reuse under CC BY 4.0 with attribution to apipriceindex.com.