Skip to content
Pricing comparison · June 2026

GLM-5.2 vs MiniMax M3: API Pricing

Input and output token rates, context windows, and real monthly cost for GLM-5.2 (Z.ai) and MiniMax M3 (MiniMax), side by side. Prices are standard on-demand rates as of June 2026.

The short answer

For a typical coding-agent workload (60M in / 12M out per month), MiniMax M3 is the cheaper option at $32.40/mo versus $137/mo for GLM-5.2 - about 76% less (4.2× cheaper). On the headline sticker of 1M input + 1M output, GLM-5.2 is $5.80 and MiniMax M3 is $1.50.

Rates at a glance

GLM-5.2MiniMax M3
Input ($/1M tokens)$1.40$0.30
Output ($/1M tokens)$4.40$1.20
Blended (1M in + 1M out)$5.80$1.50
Context window200K1,000K
TypeOpen-weightOpen-weight
ProviderZ.aiMiniMax

Monthly cost by workload

Estimated monthly API spend at each workload's token volume. Output usually costs several times input, so the winner can flip with your mix.

WorkloadGLM-5.2MiniMax M3Cheaper
Chatbot / assistant10M in / 3M out$27.20/mo$6.60/moMiniMax M3
Coding agent60M in / 12M out$137/mo$32.40/moMiniMax M3
RAG / summarization40M in / 4M out$73.60/mo$16.80/moMiniMax M3
Batch / classification20M in / 2M out$36.80/mo$8.40/moMiniMax M3

Want your own in/out split? Use the full interactive comparator to rank every model and provider for your exact workload.

Frequently asked questions

Is GLM-5.2 or MiniMax M3 cheaper?

It depends on your input/output mix, but for a typical coding-agent workload (60M in / 12M out per month) MiniMax M3 costs $32.40/mo versus $137/mo for GLM-5.2 - about 76% less (4.2x). On the headline sticker (1M input + 1M output), GLM-5.2 is $5.80 and MiniMax M3 is $1.50.

What are the token rates for GLM-5.2 and MiniMax M3?

GLM-5.2 (Z.ai) is $1.40 per 1M input and $4.40 per 1M output. MiniMax M3 (MiniMax) is $0.30 per 1M input and $1.20 per 1M output. These are standard on-demand rates, not cached or batch.

Is GLM-5.2 or MiniMax M3 open-weight?

GLM-5.2 is open-weight and MiniMax M3 is open-weight. Open-weight models can be self-hosted or run on third-party hosts at different rates, so the first-party price shown here is a starting point, not the only option.

What context window do GLM-5.2 and MiniMax M3 support?

GLM-5.2 supports 200K tokens and MiniMax M3 supports 1,000K tokens. Some models also step up pricing past a size threshold - check the source pricing pages for long-context tiers.

More pricing comparisons

Stay ahead of the AI tools curve

Picks, reviews, and automation tips every weekday. Free, no spam.