mistral-small-3 pricing
mistral-small-3 costs $0.20 per 1M input tokens and $0.80 per 1M output tokens as of 2026-09-12. A workload of 10M input and 2M output tokens per month runs $3.60. The cheapest comparable alternative is qwen-3-max at $0.60 per 1M output tokens.
mistral-small-3 is priced at $0.20 per 1M input tokens and $0.80 per 1M output tokens, with roughly 140 ms to first token and about 95 tokens/sec output throughput.
Last updated . Model pricing is refreshed twice daily.
Price per 1M tokens
Input $0.20 · Output $0.80 · Output/input ratio 4.0x.
- 1M input + 1M output tokens: $1.00
- 10M input + 2M output tokens/month: $3.60
- 100M input + 20M output tokens/month: $36.00
Cheaper alternatives to mistral-small-3
Same workload, lower output price. Validate quality before switching.
- llama-3.3-70b — $0.60/1M output (25% cheaper)
- gemini-2.5-flash — $0.60/1M output (25% cheaper)
- titan-text-express — $0.60/1M output (25% cheaper)
- qwen-3-max — $0.60/1M output (25% cheaper)
How much are you wasting on mistral-small-3?
Price lists are the easy part. Run the free waste scanner on your usage export to see which calls should have run on a cheaper tier, which prompts are oversized, and which retries you paid for twice — with dollars recoverable per finding.
Frequently asked questions
How much does mistral-small-3 cost per million tokens?
$0.20 per 1M input tokens and $0.80 per 1M output tokens, as listed on 2026-09-12.
What does mistral-small-3 cost per month at typical usage?
At 10M input and 2M output tokens per month, mistral-small-3 costs about $3.60. At 100M input and 20M output tokens it costs about $36.00.
Is there a cheaper alternative to mistral-small-3?
Yes. llama-3.3-70b at $0.60/1M output, gemini-2.5-flash at $0.60/1M output, titan-text-express at $0.60/1M output, qwen-3-max at $0.60/1M output. Validate quality on your own evaluation set before switching.
How fast is mistral-small-3?
Roughly 140 ms to first token and about 95 output tokens per second.