Loading...
Loading...
As a 30B-class SOTA model, GLM-4.7-Flash offers a new option that balances performance and efficiency. It is further optimized for agentic coding use cases, strengthening coding capabilities, long-horizon task planning,...
Best for
Price history
Local run command
# Refer to provider docs:
Presumptive Specs:
Typical GPU
12–24 GB VRAM
System RAM
16+ GB
Disk
20+ GB free
Runtime
Docker / local runtime
Community reliability estimate · not official
About this score: Community-estimated based on user reports and publicly available benchmark data (e.g. TruthfulQA). This is not an official score from the model provider. Scores may be inaccurate — always verify with the official leaderboard before making production decisions.
Input & output cost per 1M tokens over time
Price history is a Pro feature
Track pricing trends and catch price drops early.
Upgrade to Pro — $19/mo →Provider
Learn how to use Z.ai: GLM 4.7 Flash
💡 Tip: Start with the sample prompts above to see how Z.ai: GLM 4.7 Flash works best.
Proven prompts shared by the community for this model