Models/Zhipu

Zhipu

Zhipu models, pricing, and protocols

35 currently available models with pricing synced from the MoleAPI console.

35 models

Compare
Z
glm-5.3-70%

glm-5.3 is a Zhipu model available through MoleAPI with openai, openai-response, anthropic, gemini access. It fits OpenAI compatible, Responses API, Anthropic compatible, Gemini compatible workflows, with pricing synced from the MoleAPI console.

View model
defaultx1
Input$0.87 / 1M tokens
Output$3.48 / 1M tokens
Cache read$0.19 / 1M tokens
discountx0.8
Input$0.696 / 1M tokens
Output$2.784 / 1M tokens
Cache read$0.152 / 1M tokens
relayx0.3
Input$0.261 / 1M tokens
Output$1.044 / 1M tokens
Cache read$0.057 / 1M tokens
openaiopenai-responseanthropicgeminiOpenAI compatibleResponses APIAnthropic compatibleGemini compatiblePrompt cacheCompare
Z
glm-5.2-90%Official profile verified

GLM-5.2 is Z.AI's text model for long-horizon coding and engineering, with 1M context, 128K output, and configurable reasoning effort.

View model
defaultx1
Input$0.87 / 1M tokens
Output$3.48 / 1M tokens
Cache read$0.19 / 1M tokens
discountx0.8
Input$0.696 / 1M tokens
Output$2.784 / 1M tokens
Cache read$0.152 / 1M tokens
relayx0.3
Input$0.261 / 1M tokens
Output$1.044 / 1M tokens
Cache read$0.057 / 1M tokens
tempx0.1
Input$0.087 / 1M tokens
Output$0.348 / 1M tokens
Cache read$0.019 / 1M tokens
openaiopenai-responseanthropicgeminiOpenAI compatibleResponses APIAnthropic compatibleGemini compatiblePrompt cacheCompare
Z
glm-5.1-70%Official profile verified

GLM-5.1 is Z.AI's general reasoning and coding model for Chinese and English knowledge work, tools, and engineering automation.

defaultx1
Input$0.87 / 1M tokens
Output$3.48 / 1M tokens
Cache read$0.19 / 1M tokens
discountx0.8
Input$0.696 / 1M tokens
Output$2.784 / 1M tokens
Cache read$0.152 / 1M tokens
relayx0.3
Input$0.261 / 1M tokens
Output$1.044 / 1M tokens
Cache read$0.057 / 1M tokens
openaiopenai-responseanthropicgeminiOpenAI compatibleResponses APIAnthropic compatibleGemini compatiblePrompt cacheCompare
Z
glm-5-70%

GLM model by 智谱 for general chat, writing, coding assistance, and tool-assisted workflows.

defaultx1
Input$0.58 / 1M tokens
Output$2.61 / 1M tokens
Cache read$0.145 / 1M tokens
discountx0.8
Input$0.464 / 1M tokens
Output$2.088 / 1M tokens
Cache read$0.116 / 1M tokens
relayx0.3
Input$0.174 / 1M tokens
Output$0.783 / 1M tokens
Cache read$0.0435 / 1M tokens
openaiopenai-responseanthropicgeminiOpenAI compatibleResponses APIAnthropic compatibleGemini compatiblePrompt cacheCompare
Z
glm-5-turbo-70%

GLM model by 智谱 for general chat, writing, coding assistance, and tool-assisted workflows.

defaultx1
Input$0.73 / 1M tokens
Output$3.212 / 1M tokens
Cache read$0.1737 / 1M tokens
discountx0.8
Input$0.584 / 1M tokens
Output$2.5696 / 1M tokens
Cache read$0.139 / 1M tokens
relayx0.3
Input$0.219 / 1M tokens
Output$0.9636 / 1M tokens
Cache read$0.0521 / 1M tokens
openaiopenai-responseanthropicgeminiOpenAI compatibleResponses APIAnthropic compatibleGemini compatiblePrompt cacheCompare
Z
glm-4.7-70%

GLM model by 智谱 for general chat, writing, coding assistance, and tool-assisted workflows.

defaultx1
Input$0.29 / 1M tokens
Output$1.16 / 1M tokens
Cache read$0.058 / 1M tokens
Cache write$0.058 / 1M tokens
discountx0.8
Input$0.232 / 1M tokens
Output$0.928 / 1M tokens
Cache read$0.0464 / 1M tokens
Cache write$0.0464 / 1M tokens
relayx0.3
Input$0.087 / 1M tokens
Output$0.348 / 1M tokens
Cache read$0.0174 / 1M tokens
Cache write$0.0174 / 1M tokens
openaiopenai-responseanthropicgeminiOpenAI compatibleResponses APIAnthropic compatibleGemini compatiblePrompt cacheCompare
Z
glm-4.6-70%

GLM model by 智谱 for general chat, writing, coding assistance, and tool-assisted workflows.

defaultx1
Input$0.508 / 1M tokens
Output$2.032 / 1M tokens
Cache read$0.1016 / 1M tokens
discountx0.8
Input$0.4064 / 1M tokens
Output$1.6256 / 1M tokens
Cache read$0.0813 / 1M tokens
relayx0.3
Input$0.1524 / 1M tokens
Output$0.6096 / 1M tokens
Cache read$0.0305 / 1M tokens
openaiopenai-responseanthropicgeminiOpenAI compatibleResponses APIAnthropic compatibleGemini compatiblePrompt cacheCompare
Z
glm-4.5-70%

GLM model by 智谱 for general chat, writing, coding assistance, and tool-assisted workflows.

defaultx1
Input$0.67 / 1M tokens
Output$2.68 / 1M tokens
Cache read$0.134 / 1M tokens
discountx0.8
Input$0.536 / 1M tokens
Output$2.144 / 1M tokens
Cache read$0.1072 / 1M tokens
relayx0.3
Input$0.201 / 1M tokens
Output$0.804 / 1M tokens
Cache read$0.0402 / 1M tokens
openaiopenai-responseanthropicgeminiOpenAI compatibleResponses APIAnthropic compatibleGemini compatiblePrompt cacheCompare
Z
glm-4.5-air-70%

GLM model by 智谱 for general chat, writing, coding assistance, and tool-assisted workflows.

View model
defaultx1
Input$0.116 / 1M tokens
Output$0.29 / 1M tokens
Cache read$0.0232 / 1M tokens
Cache write$0.0232 / 1M tokens
discountx0.8
Input$0.0928 / 1M tokens
Output$0.232 / 1M tokens
Cache read$0.0186 / 1M tokens
Cache write$0.0186 / 1M tokens
relayx0.3
Input$0.0348 / 1M tokens
Output$0.087 / 1M tokens
Cache read$0.007 / 1M tokens
Cache write$0.007 / 1M tokens
openaiopenai-responseanthropicgeminiOpenAI compatibleResponses APIAnthropic compatibleGemini compatiblePrompt cacheCompare
Z
z-ai/glm-5.2Official profile verified

GLM-5.2 is Z.AI's text model for long-horizon coding and engineering, with 1M context, 128K output, and configurable reasoning effort.

defaultx1
Input$0.87 / 1M tokens
Output$3.48 / 1M tokens
Cache read$0.19 / 1M tokens
openaiOpenAI compatibleCompare
Z

GLM model by 智谱 for general chat, writing, coding assistance, and tool-assisted workflows.

defaultx1
Input$0.73 / 1M tokens
Output$3.2 / 1M tokens
Cache read$0.17 / 1M tokens
openaiOpenAI compatiblePrompt cacheCompare
Z

GLM model by 智谱 for general chat, writing, coding assistance, and tool-assisted workflows.

defaultx1
Input$0.09 / 1M tokens
Output$0.53 / 1M tokens
openaiOpenAI compatibleCompare
Z

GLM model by 智谱 for general chat, writing, coding assistance, and tool-assisted workflows.

View model
defaultx1
Input$0 / 1M tokens
Output$0 / 1M tokens
openaiOpenAI compatibleCompare
Z

GLM model by 智谱 for general chat, writing, coding assistance, and tool-assisted workflows.

defaultx1
Input$0.072 / 1M tokens
Output$0.432 / 1M tokens
Cache read$0.0144 / 1M tokens
Cache write$0.0144 / 1M tokens
openaiOpenAI compatiblePrompt cacheCompare
Z

GLM model by 智谱 for general chat, writing, coding assistance, and tool-assisted workflows.

defaultx1
Input$0.144 / 1M tokens
Output$0.432 / 1M tokens
Cache read$0.0288 / 1M tokens
Cache write$0.0288 / 1M tokens
openaiOpenAI compatiblePrompt cacheCompare
Z

GLM model by 智谱 for general chat, writing, coding assistance, and tool-assisted workflows.

defaultx1
Input$0.022 / 1M tokens
Output$0.2204 / 1M tokens
openaiOpenAI compatibleCompare
Z

GLM model by 智谱 for general chat, writing, coding assistance, and tool-assisted workflows.

defaultx1
Input$0.2816 / 1M tokens
Output$0.8448 / 1M tokens
openaiOpenAI compatibleCompare
Z

GLM model by 智谱 for general chat, writing, coding assistance, and tool-assisted workflows.

View model
defaultx1
Input$0 / 1M tokens
Output$0 / 1M tokens
openaiOpenAI compatibleCompare
Z

GLM model by 智谱 for general chat, writing, coding assistance, and tool-assisted workflows.

defaultx1
Input$0.5634 / 1M tokens
Output$1.1268 / 1M tokens
openaiOpenAI compatibleCompare
Z

GLM model by 智谱 for general chat, writing, coding assistance, and tool-assisted workflows.

defaultx1
Input$0.29 / 1M tokens
Output$0.87 / 1M tokens
Cache write$0.058 / 1M tokens
openaiOpenAI compatiblePrompt cacheCompare
Z
glm-4

GLM model by 智谱 for general chat, writing, coding assistance, and tool-assisted workflows.

defaultx1
Input$14.492 / 1M tokens
Output$14.492 / 1M tokens
openaiOpenAI compatibleCompare
Z

GLM model by 智谱 for general chat, writing, coding assistance, and tool-assisted workflows.

defaultx1
Input$0.274 / 1M tokens
Output$0.274 / 1M tokens
openaiOpenAI compatibleCompare
Z

GLM model by 智谱 for general chat, writing, coding assistance, and tool-assisted workflows.

defaultx1
Input$0.3 / 1M tokens
Output$1.2 / 1M tokens
openaiOpenAI compatibleCompare
Z

GLM model by 智谱 for general chat, writing, coding assistance, and tool-assisted workflows.

defaultx1
Input$0 / 1M tokens
Output$0 / 1M tokens
openaiOpenAI compatibleCompare
Z

GLM model by 智谱 for general chat, writing, coding assistance, and tool-assisted workflows.

defaultx1
Input$0.072 / 1M tokens
Output$0.072 / 1M tokens
openaiOpenAI compatibleCompare
Z

GLM model by 智谱 for general chat, writing, coding assistance, and tool-assisted workflows.

defaultx1
Input$1.45 / 1M tokens
Output$1.45 / 1M tokens
openaiOpenAI compatibleCompare
Z

GLM model by 智谱 for general chat, writing, coding assistance, and tool-assisted workflows.

defaultx1
Input$14.0845 / 1M tokens
Output$14.0845 / 1M tokens
openaiOpenAI compatibleCompare
Z

GLM model by 智谱 for general chat, writing, coding assistance, and tool-assisted workflows.

defaultx1
Input$0 / 1M tokens
Output$0 / 1M tokens
openaiOpenAI compatibleCompare
Z

GLM model by 智谱 for general chat, writing, coding assistance, and tool-assisted workflows.

defaultx1
Input$0.144 / 1M tokens
Output$0.144 / 1M tokens
openaiOpenAI compatibleCompare
Z

GLM model by 智谱 for general chat, writing, coding assistance, and tool-assisted workflows.

defaultx1
Input$0.724 / 1M tokens
Output$0.724 / 1M tokens
openaiOpenAI compatibleCompare
Z
glm-4v

GLM model by 智谱 for general chat, writing, coding assistance, and tool-assisted workflows.

defaultx1
Input$7.246 / 1M tokens
Output$7.246 / 1M tokens
openaiOpenAI compatibleVisionCompare
Z

GLM model by 智谱 for general chat, writing, coding assistance, and tool-assisted workflows.

defaultx1
Input$0.58 / 1M tokens
Output$0.58 / 1M tokens
openaiOpenAI compatibleVisionCompare
Z

GLM model by 智谱 for general chat, writing, coding assistance, and tool-assisted workflows.

defaultx1
Input$0.7144 / 1M tokens
Output$0.7144 / 1M tokens
openaiOpenAI compatibleCompare
Z

GLM model by 智谱 for general chat, writing, coding assistance, and tool-assisted workflows.

defaultx1
Input$0.144 / 1M tokens
Output$0.576 / 1M tokens
openaiOpenAI compatibleCompare
Z

GLM model by 智谱 for general chat, writing, coding assistance, and tool-assisted workflows.

defaultx1
Input$0 / 1M tokens
Output$0 / 1M tokens
openaiOpenAI compatibleCompare