GLM 4.7 Flashzhipu/glm-4-7-flash
As a 30B-class SOTA model, GLM-4.7-Flash offers a new option that balances performance and efficiency. It is further optimized for agentic coding use cases, strengthening coding capabilities, long-horizon task planning,...
Pricing across providers
| Provider | Input /1M | Output /1M | Blended /1M | Latency p50 | Format | Freshness | Action |
|---|
1 hosting hidden — prices older than 60 days are pending re-verification.
Affiliate disclosure: We may earn a commission from qualified signups. Pricing independence is enforced at the data layer — see our Editorial Independence Policy.
Works with
Point any of these clients at a hosting's base URL — they all speak at least one of this model's endpoint protocols (OPENAI_COMPATIBLE).
Capabilities
Code samples
Example using Zhipu AI — the cheapest hosting for this model as of last verification. Swap base_url and model to use a different provider from the matrix above.
from openai import OpenAI
client = OpenAI(
api_key="YOUR_API_KEY",
base_url="https://open.bigmodel.cn/api/paas/v4",
)
response = client.chat.completions.create(
model="glm-4.7-flash",
messages=[{"role": "user", "content": "Hello!"}],
)
print(response.choices[0].message.content)
Technical specs
- Context
- 203K
- Max output
- 16K
- Parameters
- —
- Release
- —
- Training cutoff
- —
- License
- —
Similar models
Compare with
- GLM 4.7 Flash vs DeepSeek V3Comparison planned — not yet published
- GLM 4.7 Flash vs DeepSeek V3 0324Comparison planned — not yet published
- GLM 4.7 Flash vs DeepSeek V3.1Comparison planned — not yet published
Frequently asked
How much does GLM 4.7 Flash cost?+−
How do I access GLM 4.7 Flash from outside China?+−
Is GLM 4.7 Flash open-source? Can I fine-tune it?+−
Is GLM 4.7 Flash OpenAI-compatible?+−
openai SDK client at the Provider's base_url and use the Provider's model name. See the Code Samples above for a copy-pasteable example.