# gemini-3-flash pricing

> gemini-3-flash costs $0.15 per 1M input tokens and $0.40 per 1M output tokens. Compare against cheaper alternatives and estimate your monthly bill — free, no login.

Source: https://tokenomy.ai/models/gemini-3-flash
Last updated: 2026-09-12
Publisher: Tokenomy — FinOps for AI
License: free to quote with attribution and a link to the source URL.

## Answer

gemini-3-flash costs $0.15 per 1M input tokens and $0.40 per 1M output tokens as of 2026-09-12. A workload of 10M input and 2M output tokens per month runs $2.30. The cheapest comparable alternative is deepseek-v4 at $0.22 per 1M output tokens.

gemini-3-flash is priced at $0.15 per 1M input tokens and $0.40 per 1M output tokens, with roughly 70 ms to first token and about 200 tokens/sec output throughput.

## Price per 1M tokens

Input $0.15 · Output $0.40 · Output/input ratio 2.7x.

- 1M input + 1M output tokens: $0.55
- 10M input + 2M output tokens/month: $2.30
- 100M input + 20M output tokens/month: $23.00

## Cheaper alternatives to gemini-3-flash

Same workload, lower output price. Validate quality before switching.

- deepseek-v3 — $0.30/1M output (25% cheaper)
- qwen-3-plus — $0.30/1M output (25% cheaper)
- ernie-lite — $0.30/1M output (25% cheaper)
- deepseek-v4 — $0.22/1M output (45% cheaper)

## How much are you wasting on gemini-3-flash?

Price lists are the easy part. Run the free waste scanner on your usage export to see which calls should have run on a cheaper tier, which prompts are oversized, and which retries you paid for twice — with dollars recoverable per finding.

## Frequently asked questions

### How much does gemini-3-flash cost per million tokens?

$0.15 per 1M input tokens and $0.40 per 1M output tokens, as listed on 2026-09-12.

### What does gemini-3-flash cost per month at typical usage?

At 10M input and 2M output tokens per month, gemini-3-flash costs about $2.30. At 100M input and 20M output tokens it costs about $23.00.

### Is there a cheaper alternative to gemini-3-flash?

Yes. deepseek-v3 at $0.30/1M output, qwen-3-plus at $0.30/1M output, ernie-lite at $0.30/1M output, deepseek-v4 at $0.22/1M output. Validate quality on your own evaluation set before switching.

### How fast is gemini-3-flash?

Roughly 70 ms to first token and about 200 output tokens per second.

## Related pages

- [Scan your AI spend free](https://tokenomy.ai/waste)
- [All model pricing](https://tokenomy.ai/models)
- [deepseek-v3 pricing](https://tokenomy.ai/models/deepseek-v3)
- [qwen-3-plus pricing](https://tokenomy.ai/models/qwen-3-plus)
- [ernie-lite pricing](https://tokenomy.ai/models/ernie-lite)
- [Token cost calculator](https://tokenomy.ai/tools/token-calculator)
