# glm-5.2 pricing

> glm-5.2 costs $0.35 per 1M input tokens and $1.40 per 1M output tokens. Compare against cheaper alternatives and estimate your monthly bill — free, no login.

Source: https://tokenomy.ai/models/glm-5-2
Last updated: 2026-09-12
Publisher: Tokenomy — FinOps for AI
License: free to quote with attribution and a link to the source URL.

## Answer

glm-5.2 costs $0.35 per 1M input tokens and $1.40 per 1M output tokens as of 2026-09-12. A workload of 10M input and 2M output tokens per month runs $6.30. The cheapest comparable alternative is llama-4-behemoth at $1.00 per 1M output tokens.

glm-5.2 is priced at $0.35 per 1M input tokens and $1.40 per 1M output tokens, with roughly 180 ms to first token and about 120 tokens/sec output throughput.

## Price per 1M tokens

Input $0.35 · Output $1.40 · Output/input ratio 4.0x.

- 1M input + 1M output tokens: $1.75
- 10M input + 2M output tokens/month: $6.30
- 100M input + 20M output tokens/month: $63.00

## Cheaper alternatives to glm-5.2

Same workload, lower output price. Validate quality before switching.

- claude-3-haiku — $1.25/1M output (11% cheaper)
- gpt-5-nano — $1.20/1M output (14% cheaper)
- claude-haiku-4 — $1.00/1M output (29% cheaper)
- llama-4-behemoth — $1.00/1M output (29% cheaper)

## How much are you wasting on glm-5.2?

Price lists are the easy part. Run the free waste scanner on your usage export to see which calls should have run on a cheaper tier, which prompts are oversized, and which retries you paid for twice — with dollars recoverable per finding.

## Frequently asked questions

### How much does glm-5.2 cost per million tokens?

$0.35 per 1M input tokens and $1.40 per 1M output tokens, as listed on 2026-09-12.

### What does glm-5.2 cost per month at typical usage?

At 10M input and 2M output tokens per month, glm-5.2 costs about $6.30. At 100M input and 20M output tokens it costs about $63.00.

### Is there a cheaper alternative to glm-5.2?

Yes. claude-3-haiku at $1.25/1M output, gpt-5-nano at $1.20/1M output, claude-haiku-4 at $1.00/1M output, llama-4-behemoth at $1.00/1M output. Validate quality on your own evaluation set before switching.

### How fast is glm-5.2?

Roughly 180 ms to first token and about 120 output tokens per second.

## Related pages

- [Scan your AI spend free](https://tokenomy.ai/waste)
- [All model pricing](https://tokenomy.ai/models)
- [claude-3-haiku pricing](https://tokenomy.ai/models/claude-3-haiku)
- [gpt-5-nano pricing](https://tokenomy.ai/models/gpt-5-nano)
- [claude-haiku-4 pricing](https://tokenomy.ai/models/claude-haiku-4)
- [Token cost calculator](https://tokenomy.ai/tools/token-calculator)
