Zhipu

Z.AI

z-ai/ · 11 models

Access 11 Z.AI models through AnonRouter's privacy-first gateway, including GLM 4.6, GLM 4.7, and GLM 4.7 Flash. Compare pricing, context windows, and capabilities across Z.AI's text routes — every request anonymized, with no payload logging.

Models

11

Modalities

1

Text

From (input)

$0.07

per 1M tokens

Max context

1M

Private routes

9 / 11

not anonymous-only

Catalog by modality

11 routes
Text11

Z.AI models11

Zhipu
GLM 4.6
z-ai/glm-4.6
Private

GLM-4.6 is a large language model developed by Zhiyuan AI, featuring strong reasoning capabilities and support for multiple languages. Supports the largest context window for processing extensive text and detailed analysis.

Private|198K context|$0.43/M input|$1.75/M output
Zhipu
GLM 4.7
z-ai/glm-4.7
E2EE

GLM-4.7 is a large language model developed by Zhiyuan AI, featuring strong reasoning capabilities and support for multiple languages. Supports the largest context window for processing extensive text and detailed analysis.

E2EE|198K context|$0.55/M input|$2.65/M output
Zhipu
GLM 4.7 Flash
z-ai/glm-4.7-flash
Private

GLM-4.7-Flash is a fast inference variant of GLM-4.7, optimized for speed while maintaining strong reasoning capabilities. Ideal for applications requiring quick responses with good quality.

Private|128K context|$0.125/M input|$0.50/M output
Zhipu
GLM 4.7 Flash Heretic
z-ai/glm-4.7-flash-heretic
Private

GLM-4.7-Flash-Heretic is an uncensored experimental variant of GLM-4.7-Flash, optimized for creative freedom and unfiltered dialogue with fast inference speed.

Private|200K context|$0.07/M input|$0.40/M output
Zhipu
GLM 5
z-ai/glm-5
Private

GLM-5 is the next-generation large language model developed by Zhiyuan AI, featuring significantly enhanced reasoning capabilities, improved instruction following, and support for multiple languages. Supports large context windows for processing extensive text and detailed analysis.

Private|198K context|$1.00/M input|$3.20/M output
Zhipu
GLM 5 Turbo
z-ai/glm-5-turbo
Anonymous

GLM-5 Turbo is a fast inference model from Z.ai tuned for strong performance in agent-driven environments and production coding workflows.

Anonymous|200K context|$1.20/M input|$4.00/M output
Zhipu
GLM 5.1
z-ai/glm-5.1
E2EE

GLM-5.1 is the next-generation large language model developed by Zhiyuan AI, featuring significantly enhanced reasoning capabilities, improved instruction following, and support for multiple languages. Supports large context windows for processing extensive text and detailed analysis with fast inference speed.

E2EE|200K context|$1.54/M input|$4.84/M output
Zhipu
GLM 5.2
z-ai/glm-5.2
E2EE

GLM-5.2 is the next-generation large language model developed by Zhiyuan AI, featuring significantly enhanced reasoning capabilities, improved instruction following, and support for multiple languages. Supports large context windows for processing extensive text and detailed analysis with fast inference speed.

E2EE|1M context|$1.40/M input|$4.40/M output
Zhipu
GLM 5V Turbo
z-ai/glm-5v-turbo
Anonymous

GLM-5V-Turbo is Z.ai's first native multimodal agent foundation model, built for vision-based coding and agent-driven tasks with image, video, and text inputs.

Anonymous|200K context|$1.50/M input|$5.00/M output
Zhipu
GLM 5.3
z-ai/glm-5.3
Private

GLM 5.3 served through the Phala AI private gateway.

Private|1M context|$1.40/M input|$4.40/M output
Zhipu
GLM 5.3 Flash
z-ai/glm-5.3-flash
Private

GLM 5.3 Flash served through the Phala AI private gateway.

Private|1M context|$0.15/M input|$0.50/M output
Z.AI models — AnonRouter