GLM AI Models
Zhipu AI with bilingual proficiency, web search, and tool integration.
5 active modelsAvailable Models
$1.40 / $4.40 per 1M
$0.85 / $3.40 per 1M
$0.70 / $3.10 per 1M
$0.11 / $0.28 per 1M
$0.00 / $0.00 per 1M
About GLM Models
Usage patterns
GLM's lineup is built for agentic coding workflows. Across 5 active variants (from $0.00 to $1.40/1M input), tool-calling and code generation are available on most models. Teams running autonomous agents typically route high-volume calls to the cheaper tier and reserve the flagship for hard reasoning steps.
Pricing across the lineup
Input pricing varies 100% across the lineup ($0.00 → $1.40/1M). At 1M requests/month with ~1K input tokens, the cheapest variant costs roughly $0.00 while the flagship runs $1.40 — a 100% gap that compounds fast at scale. Output-token pricing follows the same tiering, so long-form generation amplifies the difference.
Provider notes
GLM sits in the General Purpose category. Zhipu AI with bilingual proficiency, web search, and tool integration. The 1M maximum context window (on GLM-5.2) is the largest in this lineup — enough to ingest entire codebases or long documents in one call.
Where GLM fits
Real-world fit: GLM variants are commonly used for code-completion copilots, automated refactoring pipelines, and test generation at scale. The cheaper coding variants handle bulk linting/fixing; the flagship is reserved for architectural changes and complex bug fixes.