GLM 4.7 Flash (DeepInfra) API
DeepInfra · serves glm-4.7-flash
Pricing
| Input / 1M | $0.06 |
|---|---|
| Output / 1M | $0.40 |
| Blended / 1M | $0.46 |
| Cached input / 1M | $0.01 |
| Context window | 202,752 tokens |
| Max output | 16,384 tokens |
| License | Open weights · MIT permissive |
| Knowledge cutoff | January 2026 · source |
| Intelligence index | 23.3 Artificial Analysis Intelligence Index · source |
textfunction-calling
Price verified 2026-09-05 (0 days ago) · sourcecorroborated
Sources
| Source | Input / 1M | Output / 1M | Verified |
|---|---|---|---|
| OpenRouter | $0.06 | $0.40 | 2026-09-05 |
| LiteLLM | $0.06 | $0.40 | 2026-09-03 |
| models.dev | $0.06 | $0.40 | 2026-09-04 |
Two independent sources agree within 20%.
Call it
Served via OpenRouter, pinned to DeepInfra — the provider these prices are resolved from. Set OPENROUTER_API_KEY. Provider page.
curl https://openrouter.ai/api/v1/chat/completions \
-H "Authorization: Bearer $OPENROUTER_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "z-ai/glm-4.7-flash",
"messages": [{"role": "user", "content": "Hello"}],
"provider": {"only": ["DeepInfra"], "allow_fallbacks": false}
}'from openai import OpenAI # pip install openai
client = OpenAI(
base_url="https://openrouter.ai/api/v1",
api_key="...", # OPENROUTER_API_KEY
)
resp = client.chat.completions.create(
model="z-ai/glm-4.7-flash",
messages=[{"role": "user", "content": "Hello"}],
extra_body={"provider": {"only": ["DeepInfra"], "allow_fallbacks": false}},
)
print(resp.choices[0].message.content)Price unchanged since 2026-09-03 — 1 point on record.
Use this data
Raw JSON for this page: glm-4.7-flash.json. Free to reuse under CC BY 4.0 with attribution to apipriceindex.com.