gpt-5 vs gemini-2.5-flash
On a 1,500-input, 500-output call at 10,000 calls a day, gpt-5 costs about $3,150 a month and gemini-2.5-flash about $180 — gemini-2.5-flash is roughly 94% cheaper than gpt-5. gpt-5 lists $12.00 per 1M output tokens against gemini-2.5-flash at $0.60.
gpt-5 costs $3.00/1M in and $12.00/1M out; gemini-2.5-flash costs $0.20/1M in and $0.60/1M out.
Last updated . Model pricing is refreshed twice daily.
Price per 1M tokens
gpt-5: input $3.00, output $12.00. gemini-2.5-flash: input $0.20, output $0.60.
- Monthly at 10,000 calls/day — gpt-5: $3,150, gemini-2.5-flash: $180
- Monthly at 100,000 calls/day — gpt-5: $31,500, gemini-2.5-flash: $1,800
- Time to first token — gpt-5: 250 ms, gemini-2.5-flash: 100 ms
- Throughput — gpt-5: 180 tok/s, gemini-2.5-flash: 150 tok/s
Which one to run
If your workload tolerates gemini-2.5-flash's quality, the saving is immediate. If it does not, route: send the easy calls to gemini-2.5-flash and keep gpt-5 for the hard ones.
Frequently asked questions
Is gpt-5 cheaper than gemini-2.5-flash?
gemini-2.5-flash is cheaper. At 10,000 calls a day with 1,500 input and 500 output tokens, gpt-5 costs about $3,150 a month and gemini-2.5-flash about $180.
What is the price difference between gpt-5 and gemini-2.5-flash?
About 94% on the same workload, using list prices with no caching or batch discounts.