GLM 5 Turbo

glm-5-turbo · Z.AI

GLM-5-Turbo is a foundational model deeply optimized for the OpenClaw scenario. From the training stage it has been specifically optimized for the core requirements of OpenClaw tasks, enhancing key capabilities such as tool invocation, instruction following, scheduled and persistent tasks, and long-chain execution.

API Pricing

Input$1.2 / 1M tokens
Output$3.9996 / 1M tokens
Cache read$0.24 / 1M tokens

Specifications

Context205K tokens
Modalitiestext
CapabilitiesThinking, Streaming, Tool calling, Structured outputs, Prompt caching

Frequently asked questions

What is GLM 5 Turbo?

GLM-5-Turbo is a foundational model deeply optimized for the OpenClaw scenario. From the training stage it has been specifically optimized for the core requirements of OpenClaw tasks, enhancing key capabilities such as tool invocation, instruction following, scheduled and persistent tasks, and long-chain execution.

What is the context length of GLM 5 Turbo?

GLM 5 Turbo has a 204,800 token context window.

How much does GLM 5 Turbo cost?

On AIHubMix, GLM 5 Turbo costs $1.2 per million input tokens and $3.9996 per million output tokens. Cached input reads are billed at $0.24 per million tokens.

What modalities does GLM 5 Turbo support?

GLM 5 Turbo accepts text input.

What capabilities does GLM 5 Turbo support?

GLM 5 Turbo supports thinking, tool calling, function calling and structured outputs. Per-protocol parameter support is listed in the capability table on this page.

How do I call GLM 5 Turbo via API?

GLM 5 Turbo is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to glm-5-turbo — no other code changes needed.

Who created GLM 5 Turbo?

GLM 5 Turbo is developed by Z.AI. AIHubMix aggregates it alongside models from other providers behind one API and one bill.

When was GLM 5 Turbo released?

GLM 5 Turbo was released on March 16, 2026 by Z.AI.

More models from Z.AI

See all Z.AI models →

GLM 5.3 Flashx

by Z.AI

GLM-5.3-FlashX is Z.AI’s high-speed inference model, designed for coding agents…

$0.37/1M in · $1.25/1M out
1,000,000 tokens context

GLM 5.3 Flash

by Z.AI

GLM-5.3-Flash is a high-efficiency multimodal model from Z.AI. It supports a context…

$0.1127/1M in · $0.3944/1M out
1,048,576 tokens context

GLM 5.3

by Z.AI

GLM-5.3 is Z.AI’s coding and agentic reasoning model, built for complex software…

$1.1268/1M in · $3.9438/1M out
1,048,576 tokens context

Coding GLM 5.3

by Z.AI

GLM-5.3 is Z.ai’s reasoning model for coding and agentic workflows, designed for complex…

$0.06/1M in · $0.22/1M out
1,048,576 tokens context

Coding GLM 5.3 Flash (free)

by Z.AI

coding-glm-5.3-flash-free is the open and free version of coding-glm-5.3-flash. To ensure…

1,000,000 tokens context

Coding GLM 5.3 Flash

by Z.AI

Coding GLM 5.3 Flash is a dedicated version of GLM 5.3 Flash built for AI coding and…

$0.0282/1M in · $0.0986/1M out
1,000,000 tokens context

Use GLM 5 Turbo via the AIHubMix unified API — one interface for every major LLM.