Skip to content
Pricing comparison · June 2026

GLM-5.2 vs Qwen3-Max: API Pricing

Input and output token rates, context windows, and real monthly cost for GLM-5.2 (Z.ai) and Qwen3-Max (Alibaba), side by side. Prices are standard on-demand rates as of June 2026.

The short answer

For a typical coding-agent workload (60M in / 12M out per month), GLM-5.2 is the cheaper option at $137/mo versus $144/mo for Qwen3-Max - about 5% less. On the headline sticker of 1M input + 1M output, GLM-5.2 is $5.80 and Qwen3-Max is $7.20.

Rates at a glance

GLM-5.2Qwen3-Max
Input ($/1M tokens)$1.40$1.20
Output ($/1M tokens)$4.40$6.00
Blended (1M in + 1M out)$5.80$7.20
Context window200K262K
TypeOpen-weightProprietary
ProviderZ.aiAlibaba

Monthly cost by workload

Estimated monthly API spend at each workload's token volume. Output usually costs several times input, so the winner can flip with your mix.

WorkloadGLM-5.2Qwen3-MaxCheaper
Chatbot / assistant10M in / 3M out$27.20/mo$30.00/moGLM-5.2
Coding agent60M in / 12M out$137/mo$144/moGLM-5.2
RAG / summarization40M in / 4M out$73.60/mo$72.00/moQwen3-Max
Batch / classification20M in / 2M out$36.80/mo$36.00/moQwen3-Max

Want your own in/out split? Use the full interactive comparator to rank every model and provider for your exact workload.

Frequently asked questions

Is GLM-5.2 or Qwen3-Max cheaper?

It depends on your input/output mix, but for a typical coding-agent workload (60M in / 12M out per month) GLM-5.2 costs $137/mo versus $144/mo for Qwen3-Max - about 5% less. On the headline sticker (1M input + 1M output), GLM-5.2 is $5.80 and Qwen3-Max is $7.20.

What are the token rates for GLM-5.2 and Qwen3-Max?

GLM-5.2 (Z.ai) is $1.40 per 1M input and $4.40 per 1M output. Qwen3-Max (Alibaba) is $1.20 per 1M input and $6.00 per 1M output. These are standard on-demand rates, not cached or batch.

Is GLM-5.2 or Qwen3-Max open-weight?

GLM-5.2 is open-weight and Qwen3-Max is proprietary. Open-weight models can be self-hosted or run on third-party hosts at different rates, so the first-party price shown here is a starting point, not the only option.

What context window do GLM-5.2 and Qwen3-Max support?

GLM-5.2 supports 200K tokens and Qwen3-Max supports 262K tokens. Some models also step up pricing past a size threshold - check the source pricing pages for long-context tiers.

More pricing comparisons

Stay ahead of the AI tools curve

Picks, reviews, and automation tips every weekday. Free, no spam.