glm-4.5v · Z.AI
GLM-4.5V is a vision-language foundational model designed for multimodal agent applications. Based on a mixture-of-experts (MoE) architecture, it has 106 billion parameters and 12 billion active parameters. It delivers outstanding performance in video understanding, image question answering, OCR, and document parsing, and achieves significant improvements in front-end web encoding, basic reasoning, and spatial reasoning.
GLM-4.5V is a vision-language foundational model designed for multimodal agent applications. Based on a mixture-of-experts (MoE) architecture, it has 106 billion parameters and 12 billion active parameters. It delivers outstanding performance in video understanding, image question answering, OCR, and document parsing, and achieves significant improvements in front-end web encoding, basic reasoning, and spatial reasoning.
GLM 4.5 Vision has a 64,000 token context window.
On AIHubMix, GLM 4.5 Vision costs $0.274 per million input tokens and $0.822 per million output tokens.
GLM 4.5 Vision accepts text, image and video input.
GLM 4.5 Vision is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to glm-4.5v — no other code changes needed.
GLM 4.5 Vision is developed by Z.AI. AIHubMix aggregates it alongside models from other providers behind one API and one bill.
GLM-5.3-Flash is a high-efficiency multimodal model from Z.AI. It supports a context…
GLM-5.3 is Z.AI’s coding and agentic reasoning model, built for complex software…
GLM-5.3 is Z.ai’s reasoning model for coding and agentic workflows, designed for complex…
coding-glm-5.3-flash-free is the open and free version of coding-glm-5.3-flash. To ensure…
Coding GLM 5.3 Flash is a dedicated version of GLM 5.3 Flash built for AI coding and…
coding-glm-5.3-free is the open and free version of coding-glm-5.3. To ensure stable…
Use GLM 4.5 Vision via the AIHubMix unified API — one interface for every major LLM.