GLM-4.7 Flash Pricing
Zhipu AI · Free tier GLM model. Currently in open beta — perfect for testing and prototypin
Save up to 70% on GLM-4.7 Flash API cost
Same model, same quality — pay less per token than the official API. Pay-as-you-go, no credit card required.
Other Zhipu AI Models
GLM-5.1GLM-5-TurboGLM-4.5-AirGLM-4.7 Flash Pricing Context
Pricing position within Zhipu AI
GLM-4.7 Flash is the cheapest active model in Zhipu AI's lineup at $0.00/1M input — no sibling undercuts it. The most expensive sibling costs $0.85/1M (Infinity% more). At scale, routing high-volume calls here vs the flagship saves significantly.
When GLM-4.7 Flash is worth the cost
GLM-4.7 Flash is used for general-purpose text tasks — chat, summarization, drafting, classification, and extraction. It handles the standard text-in/text-out case reliably. For specialized workloads (coding, reasoning, vision), a purpose-tuned sibling may perform better.
Cost vs sibling models
What makes GLM-4.7 Flash different from sibling models: compared to GLM-5.1 ($0.85/1M more expensive, same 128K context); GLM-5-Turbo ($0.70/1M more expensive, same 128K context); GLM-4.5-Air ($0.11/1M more expensive, same 128K context). Choose GLM-4.7 Flash when cost per token is the priority.
Real-world cost scenarios
Real-world deployments: customer support chatbots, content drafting and summarization, classification pipelines, and extraction workflows. GLM-4.7 Flash handles the standard text-in/text-out case reliably — route specialized tasks (vision, coding, reasoning) to purpose-tuned siblings.