deepseek-v4 pricing
deepseek-v4 costs $0.11 per 1M input tokens and $0.22 per 1M output tokens as of 2026-09-12. A workload of 10M input and 2M output tokens per month runs $1.54. The cheapest comparable alternative is gemini-2.5-flash-lite at $0.15 per 1M output tokens.
deepseek-v4 is priced at $0.11 per 1M input tokens and $0.22 per 1M output tokens, with roughly 180 ms to first token and about 80 tokens/sec output throughput.
Last updated . Model pricing is refreshed twice daily.
Price per 1M tokens
Input $0.11 · Output $0.22 · Output/input ratio 2.0x.
- 1M input + 1M output tokens: $0.33
- 10M input + 2M output tokens/month: $1.54
- 100M input + 20M output tokens/month: $15.40
Cheaper alternatives to deepseek-v4
Same workload, lower output price. Validate quality before switching.
- llama-4-scout — $0.20/1M output (9% cheaper)
- phi-4-mini — $0.20/1M output (9% cheaper)
- gemini-2.5-flash-lite — $0.15/1M output (32% cheaper)
How much are you wasting on deepseek-v4?
Price lists are the easy part. Run the free waste scanner on your usage export to see which calls should have run on a cheaper tier, which prompts are oversized, and which retries you paid for twice — with dollars recoverable per finding.
Frequently asked questions
How much does deepseek-v4 cost per million tokens?
$0.11 per 1M input tokens and $0.22 per 1M output tokens, as listed on 2026-09-12.
What does deepseek-v4 cost per month at typical usage?
At 10M input and 2M output tokens per month, deepseek-v4 costs about $1.54. At 100M input and 20M output tokens it costs about $15.40.
Is there a cheaper alternative to deepseek-v4?
Yes. llama-4-scout at $0.20/1M output, phi-4-mini at $0.20/1M output, gemini-2.5-flash-lite at $0.15/1M output. Validate quality on your own evaluation set before switching.
How fast is deepseek-v4?
Roughly 180 ms to first token and about 80 output tokens per second.