llama-4-maverick pricing
llama-4-maverick is priced at $0.50 per 1M input tokens and $0.50 per 1M output tokens, with roughly 280 ms to first token and about 55 tokens/sec output throughput.
Price per 1M tokens
Input $0.50 · Output $0.50 · Output/input ratio 1.0x.
- 1M input + 1M output tokens: $1.00
- 10M input + 2M output tokens/month: $6.00
- 100M input + 20M output tokens/month: $60.00
Cheaper alternatives to llama-4-maverick
Same workload, lower output price. Validate quality before switching.
- gemini-3-flash — $0.40/1M output (20% cheaper)
- phi-4 — $0.40/1M output (20% cheaper)
- deepseek-v3 — $0.30/1M output (40% cheaper)
- qwen-3-plus — $0.30/1M output (40% cheaper)
How much are you wasting on llama-4-maverick?
Price lists are the easy part. Run the free waste scanner on your usage export to see which calls should have run on a cheaper tier, which prompts are oversized, and which retries you paid for twice — with dollars recoverable per finding.