by Z.AI
GLM-4.5V is a vision-language foundational model designed for multimodal agent applications. Based on a mixture-of-experts (MoE) architecture, it has 106 billion parameters and 12 billion active parameters. It delivers outstanding performance in video understanding, image question answering, OCR, and document parsing, and achieves significant improvements in front-end web encoding, basic reasoning, and spatial reasoning.
GLM-4.5V is a vision-language foundational model designed for multimodal agent applications. Based on a mixture-of-experts (MoE) architecture, it has 106 billion parameters and 12 billion active parameters. It delivers outstanding performance in video understanding, image question answering, OCR, and document parsing, and achieves significant improvements in front-end web encoding, basic reasoning, and spatial reasoning.
glm-4.5v has a 64,000 token context window.
On AIHubMix, glm-4.5v costs $0.27 per million input tokens and $0.82 per million output tokens.
glm-4.5v accepts text, image and video input.
glm-4.5v is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to glm-4.5v — no other code changes needed.
glm-4.5v is developed by Z.AI. AIHubMix aggregates it alongside models from other providers behind one API and one bill.
GLM-5.2 is Z.ai’s flagship model for the era of long-horizon tasks. With a truly usable…
GLM-5.2-Fast-Preview is the high-speed version of Zhipu AI’s flagship model GLM-5.2…
coding-glm-5.2-free is the open and free version of coding-glm-5.2. To maintain reliable…
Currently, the special resources for this model are limited, but due to its popularity…
GLM-5.1 is Zhipu's latest flagship model, with greatly enhanced coding capabilities and…
GLM-Image is Zhipu AI's new flagship image generation model. The model is trained…
Use glm-4.5v via the AIHubMix unified API — one interface for every major LLM.