gpt-5.2 vs llama-4-maverick

On a 1,500-input, 500-output call at 10,000 calls a day, gpt-5.2 costs about $2,100 a month and llama-4-maverick about $300 — llama-4-maverick is roughly 86% cheaper than gpt-5.2. gpt-5.2 lists $8.00 per 1M output tokens against llama-4-maverick at $0.50.

gpt-5.2 costs $2.00/1M in and $8.00/1M out; llama-4-maverick costs $0.50/1M in and $0.50/1M out.

Last updated . Model pricing is refreshed twice daily.

Price per 1M tokens

gpt-5.2: input $2.00, output $8.00. llama-4-maverick: input $0.50, output $0.50.

  • Monthly at 10,000 calls/day — gpt-5.2: $2,100, llama-4-maverick: $300
  • Monthly at 100,000 calls/day — gpt-5.2: $21,000, llama-4-maverick: $3,000
  • Time to first token — gpt-5.2: 200 ms, llama-4-maverick: 280 ms
  • Throughput — gpt-5.2: 220 tok/s, llama-4-maverick: 55 tok/s

Which one to run

If your workload tolerates llama-4-maverick's quality, the saving is immediate. If it does not, route: send the easy calls to llama-4-maverick and keep gpt-5.2 for the hard ones.

Frequently asked questions

Is gpt-5.2 cheaper than llama-4-maverick?

llama-4-maverick is cheaper. At 10,000 calls a day with 1,500 input and 500 output tokens, gpt-5.2 costs about $2,100 a month and llama-4-maverick about $300.

What is the price difference between gpt-5.2 and llama-4-maverick?

About 86% on the same workload, using list prices with no caching or batch discounts.