gpt-5.2 vs qwen-3-max

On a 1,500-input, 500-output call at 10,000 calls a day, gpt-5.2 costs about $2,100 a month and qwen-3-max about $180 — qwen-3-max is roughly 91% cheaper than gpt-5.2. gpt-5.2 lists $8.00 per 1M output tokens against qwen-3-max at $0.60.

gpt-5.2 costs $2.00/1M in and $8.00/1M out; qwen-3-max costs $0.20/1M in and $0.60/1M out.

Last updated . Model pricing is refreshed twice daily.

Price per 1M tokens

gpt-5.2: input $2.00, output $8.00. qwen-3-max: input $0.20, output $0.60.

  • Monthly at 10,000 calls/day — gpt-5.2: $2,100, qwen-3-max: $180
  • Monthly at 100,000 calls/day — gpt-5.2: $21,000, qwen-3-max: $1,800
  • Time to first token — gpt-5.2: 200 ms, qwen-3-max: 350 ms
  • Throughput — gpt-5.2: 220 tok/s, qwen-3-max: 35 tok/s

Which one to run

If your workload tolerates qwen-3-max's quality, the saving is immediate. If it does not, route: send the easy calls to qwen-3-max and keep gpt-5.2 for the hard ones.

Frequently asked questions

Is gpt-5.2 cheaper than qwen-3-max?

qwen-3-max is cheaper. At 10,000 calls a day with 1,500 input and 500 output tokens, gpt-5.2 costs about $2,100 a month and qwen-3-max about $180.

What is the price difference between gpt-5.2 and qwen-3-max?

About 91% on the same workload, using list prices with no caching or batch discounts.