# gpt-5 vs gemini-2.5-flash

> gpt-5 vs gemini-2.5-flash: cost per 1M tokens, latency, throughput and monthly bill. gemini-2.5-flash runs about 94% cheaper on the same workload.

Source: https://tokenomy.ai/compare/gpt-5-vs-gemini-2-5-flash
Last updated: 2026-09-21
Publisher: Tokenomy — FinOps for AI
License: free to quote with attribution and a link to the source URL.

## Answer

On a 1,500-input, 500-output call at 10,000 calls a day, gpt-5 costs about $3,150 a month and gemini-2.5-flash about $180 — gemini-2.5-flash is roughly 94% cheaper than gpt-5. gpt-5 lists $12.00 per 1M output tokens against gemini-2.5-flash at $0.60.

gpt-5 costs $3.00/1M in and $12.00/1M out; gemini-2.5-flash costs $0.20/1M in and $0.60/1M out.

## Price per 1M tokens

gpt-5: input $3.00, output $12.00. gemini-2.5-flash: input $0.20, output $0.60.

- Monthly at 10,000 calls/day — gpt-5: $3,150, gemini-2.5-flash: $180
- Monthly at 100,000 calls/day — gpt-5: $31,500, gemini-2.5-flash: $1,800
- Time to first token — gpt-5: 250 ms, gemini-2.5-flash: 100 ms
- Throughput — gpt-5: 180 tok/s, gemini-2.5-flash: 150 tok/s

## Which one to run

If your workload tolerates gemini-2.5-flash's quality, the saving is immediate. If it does not, route: send the easy calls to gemini-2.5-flash and keep gpt-5 for the hard ones.

## Frequently asked questions

### Is gpt-5 cheaper than gemini-2.5-flash?

gemini-2.5-flash is cheaper. At 10,000 calls a day with 1,500 input and 500 output tokens, gpt-5 costs about $3,150 a month and gemini-2.5-flash about $180.

### What is the price difference between gpt-5 and gemini-2.5-flash?

About 94% on the same workload, using list prices with no caching or batch discounts.

## Related pages

- [Scan your AI spend free](https://tokenomy.ai/waste)
- [All comparisons](https://tokenomy.ai/compare)
- [gpt-5 pricing](https://tokenomy.ai/models/gpt-5)
- [gemini-2.5-flash pricing](https://tokenomy.ai/models/gemini-2-5-flash)
- [Same model, different hosts](https://tokenomy.ai/endpoints)
