gemini-3.5-pro pricing

gemini-3.5-pro costs $2.50 per 1M input tokens and $12.00 per 1M output tokens as of 2026-09-12. A workload of 10M input and 2M output tokens per month runs $49.00. The cheapest comparable alternative is gpt-4.1 at $8.00 per 1M output tokens.

gemini-3.5-pro is priced at $2.50 per 1M input tokens and $12.00 per 1M output tokens, with roughly 140 ms to first token and about 140 tokens/sec output throughput.

Last updated . Model pricing is refreshed twice daily.

Price per 1M tokens

Input $2.50 · Output $12.00 · Output/input ratio 4.8x.

  • 1M input + 1M output tokens: $14.50
  • 10M input + 2M output tokens/month: $49.00
  • 100M input + 20M output tokens/month: $490.00

Cheaper alternatives to gemini-3.5-pro

Same workload, lower output price. Validate quality before switching.

  • gemini-2.5-pro — $10.00/1M output (17% cheaper)
  • llama-4.1-405b — $8.00/1M output (33% cheaper)
  • gpt-5.2 — $8.00/1M output (33% cheaper)
  • gpt-4.1 — $8.00/1M output (33% cheaper)

How much are you wasting on gemini-3.5-pro?

Price lists are the easy part. Run the free waste scanner on your usage export to see which calls should have run on a cheaper tier, which prompts are oversized, and which retries you paid for twice — with dollars recoverable per finding.

Frequently asked questions

How much does gemini-3.5-pro cost per million tokens?

$2.50 per 1M input tokens and $12.00 per 1M output tokens, as listed on 2026-09-12.

What does gemini-3.5-pro cost per month at typical usage?

At 10M input and 2M output tokens per month, gemini-3.5-pro costs about $49.00. At 100M input and 20M output tokens it costs about $490.00.

Is there a cheaper alternative to gemini-3.5-pro?

Yes. gemini-2.5-pro at $10.00/1M output, llama-4.1-405b at $8.00/1M output, gpt-5.2 at $8.00/1M output, gpt-4.1 at $8.00/1M output. Validate quality on your own evaluation set before switching.

How fast is gemini-3.5-pro?

Roughly 140 ms to first token and about 140 output tokens per second.