deepinfra
V4 Flash
Canonical id: deepinfra/deepinfra-deepseek-ai-DeepSeek-V4-Flash · Context: 1M tokens
Text-only
model=auto
Price per 1M tokens
Input $0.09/M · Output $0.18/M
+ LLMRouter platform fee: 1% of provider cost (≈ $0.0009/M in · $0.0018/M out) — billed as its own line, never blended into the price above.
Providers
| Provider | Context | Input $/M | Output $/M |
|---|---|---|---|
| deepinfra (viewing) | 1M | $0.09/M | $0.18/M |
| siliconflow | 1M | $0.13/M | $0.28/M |
| deepseek | 1M | $0.14/M | $0.28/M |
| fireworks | 1M | $0.14/M | $0.28/M |
| gmicloud | 1M | $0.14/M | $0.28/M |
| novita | 1M | $0.14/M | $0.28/M |
| qwen | 1M | $0.2/M | $0.4/M |
| qwen-eu | 1M | $0.2/M | $0.4/M |
Served through deepinfra's API — 0% additional markup beyond our 1% platform fee. Other rows above are the same model through a different provider — click one to switch.
Capability profile
Reasoning●●●●○4/5
Coding●●●●○4/5
Quality●●●●○4/5
Derived from Artificial Analysis's licensed Intelligence/Coding Index — see the leaderboard for their full, independently measured scores. Source: Artificial Analysis.
API
curl https://api.llmrouter.sh/v1/chat/completions \
-H "Authorization: Bearer llmr_sk_..." \
-H "Content-Type: application/json" \
-d '{"model": "deepinfra-deepseek-ai-DeepSeek-V4-Flash", "messages": [{"role": "user", "content": "Hello!"}]}'