glm-4-flash

by 智谱 ChatGLM

API Pricing

Input$0.1 / 1M tokens
Output$0.1 / 1M tokens

FAQ

How much does glm-4-flash cost?

On AIHubMix, glm-4-flash costs $0.1 per million input tokens and $0.1 per million output tokens.

How do I call glm-4-flash via API?

glm-4-flash is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to glm-4-flash — no other code changes needed.

Who develops glm-4-flash?

glm-4-flash is developed by 智谱 ChatGLM. AIHubMix aggregates it alongside models from other providers behind one API and one bill.

More from 智谱 ChatGLM

embedding-2

by 智谱 ChatGLM

A text vector model that converts input text information into vector representations so…

$0.07/1M in · $0.07/1M out
8,000 tokens context

embedding-3

by 智谱 ChatGLM

A text vector model that converts input text into vector representations to work with a…

$0.07/1M in · $0.07/1M out
8,000 tokens context

chatglm_lite

by 智谱 ChatGLM
$0.29/1M in · $0.29/1M out

chatglm_pro

by 智谱 ChatGLM
$1.43/1M in · $1.43/1M out

chatglm_std

by 智谱 ChatGLM
$0.71/1M in · $0.71/1M out

chatglm_turbo

by 智谱 ChatGLM
$0.71/1M in · $0.71/1M out

Use glm-4-flash via the AIHubMix unified API — one interface for every major LLM.