Coding GLM 5.3 Flash

coding-glm-5.3-flash · Z.AI

Coding GLM 5.3 Flash is a dedicated version of GLM 5.3 Flash built for AI coding and Coding Agent workflows. It is designed for code understanding, generation, editing, repository-level development, and automated software engineering tasks. With support for context windows of up to approximately 1 million tokens, it can handle large codebases and extended development sessions. The model also supports text, image, and video inputs, along with tool use, making it well suited for AI coding tools such as Claude Code, OpenCode, Cline, and other agentic development environments.

API Pricing

Input$0.0282 / 1M tokens
Output$0.0986 / 1M tokens
Cache read$0.007 / 1M tokens

Specifications

Context1M tokens
Modalitiestext, image, video
CapabilitiesThinking, Streaming, Tool calling, Prompt caching

Frequently asked questions

What is GLM 5.3 Flash (coding)?

Coding GLM 5.3 Flash is a dedicated version of GLM 5.3 Flash built for AI coding and Coding Agent workflows. It is designed for code understanding, generation, editing, repository-level development, and automated software engineering tasks. With support for context windows of up to approximately 1 million tokens, it can handle large codebases and extended development sessions. The model also supports text, image, and video inputs, along with tool use, making it well suited for AI coding tools such as Claude Code, OpenCode, Cline, and other agentic development environments.

What is the context length of GLM 5.3 Flash (coding)?

GLM 5.3 Flash (coding) has a 1,000,000 token context window.

How much does GLM 5.3 Flash (coding) cost?

On AIHubMix, GLM 5.3 Flash (coding) costs $0.0282 per million input tokens and $0.0986 per million output tokens. Cached input reads are billed at $0.007 per million tokens.

What modalities does GLM 5.3 Flash (coding) support?

GLM 5.3 Flash (coding) accepts text, image and video input.

What capabilities does GLM 5.3 Flash (coding) support?

GLM 5.3 Flash (coding) supports thinking, tool calling, function calling and structured outputs. Per-protocol parameter support is listed in the capability table on this page.

How do I call GLM 5.3 Flash (coding) via API?

GLM 5.3 Flash (coding) is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to coding-glm-5.3-flash — no other code changes needed.

Who created GLM 5.3 Flash (coding)?

GLM 5.3 Flash (coding) is developed by Z.AI. AIHubMix aggregates it alongside models from other providers behind one API and one bill.

When was GLM 5.3 Flash (coding) released?

GLM 5.3 Flash (coding) was released on August 26, 2026 by Z.AI.

Free version: Coding GLM 5.3 Flash (free)

More models from Z.AI

See all Z.AI models →

GLM 5.3 Flashx

by Z.AI

GLM-5.3-FlashX is Z.AI’s high-speed inference model, designed for coding agents…

$0.37/1M in · $1.25/1M out
1,000,000 tokens context

GLM 5.3 Flash

by Z.AI

GLM-5.3-Flash is a high-efficiency multimodal model from Z.AI. It supports a context…

$0.1127/1M in · $0.3944/1M out
1,048,576 tokens context

GLM 5.3

by Z.AI

GLM-5.3 is Z.AI’s coding and agentic reasoning model, built for complex software…

$1.1268/1M in · $3.9438/1M out
1,048,576 tokens context

Coding GLM 5.3

by Z.AI

GLM-5.3 is Z.ai’s reasoning model for coding and agentic workflows, designed for complex…

$0.06/1M in · $0.22/1M out
1,048,576 tokens context

Coding GLM 5.3 Flash (free)

by Z.AI

coding-glm-5.3-flash-free is the open and free version of coding-glm-5.3-flash. To ensure…

1,000,000 tokens context

Coding GLM 5.3 (free)

by Z.AI

coding-glm-5.3-free is the open and free version of coding-glm-5.3. To ensure stable…

1,048,576 tokens context

Use Coding GLM 5.3 Flash via the AIHubMix unified API — one interface for every major LLM.