gpt-4.1-mini pricing

gpt-4.1-mini costs $0.50 per 1M input tokens and $2.00 per 1M output tokens as of 2026-09-12. A workload of 10M input and 2M output tokens per month runs $9.00. The cheapest comparable alternative is gpt-5-nano at $1.20 per 1M output tokens.

gpt-4.1-mini is priced at $0.50 per 1M input tokens and $2.00 per 1M output tokens, with roughly 150 ms to first token and about 110 tokens/sec output throughput.

Last updated . Model pricing is refreshed twice daily.

Price per 1M tokens

Input $0.50 · Output $2.00 · Output/input ratio 4.0x.

  • 1M input + 1M output tokens: $2.50
  • 10M input + 2M output tokens/month: $9.00
  • 100M input + 20M output tokens/month: $90.00

Cheaper alternatives to gpt-4.1-mini

Same workload, lower output price. Validate quality before switching.

  • gpt-4o-mini — $1.50/1M output (25% cheaper)
  • glm-5.2 — $1.40/1M output (30% cheaper)
  • claude-3-haiku — $1.25/1M output (38% cheaper)
  • gpt-5-nano — $1.20/1M output (40% cheaper)

How much are you wasting on gpt-4.1-mini?

Price lists are the easy part. Run the free waste scanner on your usage export to see which calls should have run on a cheaper tier, which prompts are oversized, and which retries you paid for twice — with dollars recoverable per finding.

Frequently asked questions

How much does gpt-4.1-mini cost per million tokens?

$0.50 per 1M input tokens and $2.00 per 1M output tokens, as listed on 2026-09-12.

What does gpt-4.1-mini cost per month at typical usage?

At 10M input and 2M output tokens per month, gpt-4.1-mini costs about $9.00. At 100M input and 20M output tokens it costs about $90.00.

Is there a cheaper alternative to gpt-4.1-mini?

Yes. gpt-4o-mini at $1.50/1M output, glm-5.2 at $1.40/1M output, claude-3-haiku at $1.25/1M output, gpt-5-nano at $1.20/1M output. Validate quality on your own evaluation set before switching.

How fast is gpt-4.1-mini?

Roughly 150 ms to first token and about 110 output tokens per second.