Coding GLM 5.3 Flash

coding-glm-5.3-flash · Z.AI

Coding GLM 5.3 Flash is a dedicated version of GLM 5.3 Flash built for AI coding and Coding Agent workflows. It is designed for code understanding, generation, editing, repository-level development, and automated software engineering tasks. With support for context windows of up to approximately 1 million tokens, it can handle large codebases and extended development sessions. The model also supports text, image, and video inputs, along with tool use, making it well suited for AI coding tools such as Claude Code, OpenCode, Cline, and other agentic development environments.

API Pricing

Input$0.028 / 1M tokens
Output$0.099 / 1M tokens
Cache read$0.007 / 1M tokens

Specifications

Context1M tokens
Featuresthinking, tool calling, function calling, structured outputs

Frequently asked questions

What is Coding GLM 5.3 Flash?

Coding GLM 5.3 Flash is a dedicated version of GLM 5.3 Flash built for AI coding and Coding Agent workflows. It is designed for code understanding, generation, editing, repository-level development, and automated software engineering tasks. With support for context windows of up to approximately 1 million tokens, it can handle large codebases and extended development sessions. The model also supports text, image, and video inputs, along with tool use, making it well suited for AI coding tools such as Claude Code, OpenCode, Cline, and other agentic development environments.

What is the context length of Coding GLM 5.3 Flash?

Coding GLM 5.3 Flash has a 1,000,000 token context window.

How much does Coding GLM 5.3 Flash cost?

On AIHubMix, Coding GLM 5.3 Flash costs $0.028 per million input tokens and $0.099 per million output tokens. Cached input reads are billed at $0.007 per million tokens.

What capabilities does Coding GLM 5.3 Flash support?

Coding GLM 5.3 Flash supports thinking, tool calling, function calling and structured outputs. Per-protocol parameter support is listed in the capability table on this page.

How do I call Coding GLM 5.3 Flash via API?

Coding GLM 5.3 Flash is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to coding-glm-5.3-flash — no other code changes needed.

Who created Coding GLM 5.3 Flash?

Coding GLM 5.3 Flash is developed by Z.AI. AIHubMix aggregates it alongside models from other providers behind one API and one bill.

Free version: Coding GLM 5.3 Flash (free)

More models from Z.AI

See all Z.AI models →

GLM 5.3 Flash

by Z.AI

GLM-5.3-Flash is a high-efficiency multimodal model from Z.AI. It supports a context…

$0.113 $0.056/1M in · $0.394 $0.197/1M out
50% off
1,000,000 tokens context

GLM 5.3

by Z.AI

GLM-5.3 is Z.AI’s coding and agentic reasoning model, built for complex software…

$1.127 $1.014/1M in · $3.944 $3.549/1M out
10% off
1,000,000 tokens context

Coding GLM 5.3

by Z.AI

GLM-5.3 is Z.ai’s reasoning model for coding and agentic workflows, designed for complex…

$0.06/1M in · $0.22/1M out

Coding GLM 5.3 Flash (free)

by Z.AI

coding-glm-5.3-flash-free is the open and free version of coding-glm-5.3-flash. To ensure…

Coding GLM 5.3 (free)

by Z.AI

coding-glm-5.3-free is the open and free version of coding-glm-5.3. To ensure stable…

Ox Alpha

by Z.AI

This model actually points to glm-5.3-flash; if you need to use it in production, you can…

1,048,576 tokens context

Use Coding GLM 5.3 Flash via the AIHubMix unified API — one interface for every major LLM.