G

GLM Pricing

Zhipu AI with bilingual proficiency, web search, and tool integration.

Cost optimization across the full model lineup

Every model below is available at up to 70% lower cost than official pricing. Pay-as-you-go, no credit card.

Up to 70%
API cost savings

All GLM Models

GLM-5.2
$1.40 / $4.40
View Pricing
GLM-5.1
$0.85 / $3.40
View Pricing
GLM-5-Turbo
$0.70 / $3.10
View Pricing
GLM-4.5-Air
$0.11 / $0.28
View Pricing
GLM-4.7 Flash
$0.00 / $0.00
View Pricing

GLM Pricing Guide

How GLM pricing scales

Input pricing varies 100% across the lineup ($0.00 → $1.40/1M). At 1M requests/month with ~1K input tokens, the cheapest variant costs roughly $0.00 while the flagship runs $1.40 — a 100% gap that compounds fast at scale. Output-token pricing follows the same tiering, so long-form generation amplifies the difference.

Which variant to pick

GLM's lineup is built for agentic coding workflows. Across 5 active variants (from $0.00 to $1.40/1M input), tool-calling and code generation are available on most models. Teams running autonomous agents typically route high-volume calls to the cheaper tier and reserve the flagship for hard reasoning steps.

GLM notes

GLM sits in the General Purpose category. Zhipu AI with bilingual proficiency, web search, and tool integration. The 1M maximum context window (on GLM-5.2) is the largest in this lineup — enough to ingest entire codebases or long documents in one call.

Cost in practice

Real-world fit: GLM variants are commonly used for code-completion copilots, automated refactoring pipelines, and test generation at scale. The cheaper coding variants handle bulk linting/fixing; the flagship is reserved for architectural changes and complex bug fixes.

Get GLM API Now

إنشاء حساب
احصل على مفتاح API