GLM Pricing
Zhipu AI with bilingual proficiency, web search, and tool integration.
Cost optimization across the full model lineup
Every model below is available at up to 70% lower cost than official pricing. Pay-as-you-go, no credit card.
All GLM Models
GLM-5.1GLM-5-TurboGLM-4.5-AirGLM-4.7 FlashGLM Pricing Guide
How GLM pricing scales
Input pricing varies 100% across the lineup ($0.00 → $0.85/1M). At 1M requests/month with ~1K input tokens, the cheapest variant costs roughly $0.00 while the flagship runs $0.85 — a 100% gap that compounds fast at scale. Output-token pricing follows the same tiering, so long-form generation amplifies the difference.
Which variant to pick
GLM offers 4 active variants at $0.00–$0.85/1M input. The spread lets you match cost to task complexity: light variants for high-volume classification, mid-tier for general chat, and the flagship for demanding generation.
GLM notes
GLM sits in the General Purpose category. Zhipu AI with bilingual proficiency, web search, and tool integration. The 128K maximum context window (on GLM-5.1) is the largest in this lineup — suitable for most multi-document workloads.
Cost in practice
Real-world fit: GLM variants are deployed for customer-support chat, content drafting, and classification at scale. The budget tier absorbs high-volume routine traffic; the flagship handles nuanced long-form generation.