G

GLM-4.7 Flash Pricing

Zhipu AI · Free tier GLM model. Currently in open beta — perfect for testing and prototypin

TL;DR
Price: $0.00/1M input · $0.00/1M output
Context: 128K · max 4,096 output
Cost advantage: Up to 70% cheaper than official API
Access: OpenAI-compatible · No credit card · Instant API key

Save up to 70% on GLM-4.7 Flash API cost

Same model, same quality — pay less per token than the official API. Pay-as-you-go, no credit card required.

From $0.00/1M
input token price
INPUT
$0.00
/1M tokens
OUTPUT
$0.00
/1M tokens
CONTEXT
128K

Other Zhipu AI Models

GLM-5.1
$0.85 / $3.40
View Pricing
GLM-5-Turbo
$0.70 / $3.10
View Pricing
GLM-4.5-Air
$0.11 / $0.28
View Pricing

GLM-4.7 Flash Pricing Context

Pricing position within Zhipu AI

GLM-4.7 Flash is the cheapest active model in Zhipu AI's lineup at $0.00/1M input — no sibling undercuts it. The most expensive sibling costs $0.85/1M (Infinity% more). At scale, routing high-volume calls here vs the flagship saves significantly.

When GLM-4.7 Flash is worth the cost

GLM-4.7 Flash is used for general-purpose text tasks — chat, summarization, drafting, classification, and extraction. It handles the standard text-in/text-out case reliably. For specialized workloads (coding, reasoning, vision), a purpose-tuned sibling may perform better.

Cost vs sibling models

What makes GLM-4.7 Flash different from sibling models: compared to GLM-5.1 ($0.85/1M more expensive, same 128K context); GLM-5-Turbo ($0.70/1M more expensive, same 128K context); GLM-4.5-Air ($0.11/1M more expensive, same 128K context). Choose GLM-4.7 Flash when cost per token is the priority.

Real-world cost scenarios

Real-world deployments: customer support chatbots, content drafting and summarization, classification pipelines, and extraction workflows. GLM-4.7 Flash handles the standard text-in/text-out case reliably — route specialized tasks (vision, coding, reasoning) to purpose-tuned siblings.

Get GLM-4.7 Flash API Now

إنشاء حساب
احصل على مفتاح API