glm-4.7-flash-free · Z.AI
The glm-4.7-flash free model has usage restrictions to ensure stable service operation: a maximum of 5 requests per minute, no more than 500 requests per day, and a daily usage quota of 1 million tokens.
The glm-4.7-flash free model has usage restrictions to ensure stable service operation: a maximum of 5 requests per minute, no more than 500 requests per day, and a daily usage quota of 1 million tokens.
GLM 4.7 Flash (free) has a 200,000 token context window.
GLM 4.7 Flash (free) accepts text input.
GLM 4.7 Flash (free) supports thinking, tool calling, structured outputs and function calling. Per-protocol parameter support is listed in the capability table on this page.
GLM 4.7 Flash (free) is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to glm-4.7-flash-free — no other code changes needed.
GLM 4.7 Flash (free) is developed by Z.AI. AIHubMix aggregates it alongside models from other providers behind one API and one bill.
GLM 4.7 Flash (free) was released on January 19, 2026 by Z.AI.
GLM-5.3-FlashX is Z.AI’s high-speed inference model, designed for coding agents…
GLM-5.3-Flash is a high-efficiency multimodal model from Z.AI. It supports a context…
GLM-5.3 is Z.AI’s coding and agentic reasoning model, built for complex software…
GLM-5.3 is Z.ai’s reasoning model for coding and agentic workflows, designed for complex…
coding-glm-5.3-flash-free is the open and free version of coding-glm-5.3-flash. To ensure…
Coding GLM 5.3 Flash is a dedicated version of GLM 5.3 Flash built for AI coding and…
Use GLM 4.7 Flash (free) via the AIHubMix unified API — one interface for every major LLM.