deepinfra
Llama 4 Maverick 17B 128E Instruct FP8
Canonical id: deepinfra/deepinfra-meta-llama-Llama-4-Maverick-17B-128E-Instruct-FP8 · Context: 1M tokens
Vision
model=auto
Price per 1M tokens
Input $0.2/M · Output $0.8/M
+ LLMRouter platform fee: 1% of provider cost (≈ $0.002/M in · $0.008/M out) — billed as its own line, never blended into the price above.
Providers
| Provider | Context | Input $/M | Output $/M |
|---|---|---|---|
| deepinfra (viewing) | 1M | $0.2/M | $0.8/M |
| novita | 1M | $0.27/M | $0.85/M |
Served through deepinfra's API — 0% additional markup beyond our 1% platform fee. Other rows above are the same model through a different provider — click one to switch.
Capability profile
Reasoning●○○○○1/5
Coding●●○○○2/5
Quality●○○○○1/5
Derived from Artificial Analysis's licensed Intelligence/Coding Index — see the leaderboard for their full, independently measured scores. Source: Artificial Analysis.
API
curl https://api.llmrouter.sh/v1/chat/completions \
-H "Authorization: Bearer llmr_sk_..." \
-H "Content-Type: application/json" \
-d '{"model": "deepinfra-meta-llama-Llama-4-Maverick-17B-128E-Instruct-FP8", "messages": [{"role": "user", "content": "Hello!"}]}'