z-ai/glm-4.7-flash
As a 30B-class SOTA model, GLM-4.7-Flash offers a new option that balances performance and efficiency. It is further optimized for agentic coding use cases, strengthening coding capabilities, long-horizon task planning,...
The most important technical and pricing details in one place.
Gateway-compatible routes are tried from the lowest cost; other sources are shown for catalog reference.

Current request, token, and latency data across TamgaStudio.
Call Z.ai: GLM 4.7 Flash with a single request through the OpenAI-compatible TamgaStudio API. Create your API key from the API Keys section in your dashboard.
curl https://api.tamga.studio/v1/chat/completions \
-H "Content-Type: application/json" \
-H "Authorization: Bearer tl-sk-XXXXXX" \
-d '{"model": "z-ai/glm-4.7-flash", "messages": [{"role": "user", "content": "Hello!"}]}'
Common questions about this model.
Explore other models from the same provider or in a similar price range as Z.ai: GLM 4.7 Flash.