Zhipu

GLM 5 Turbo

z-ai
Chat

GLM-5 Turbo is a fast inference model from Z.ai tuned for strong performance in agent-driven environments and production coding workflows.

Modalities

In / out price

$1.20 / $4.00

per 1M

Cached price

$0.24

per 1M

Context

200K

Max output

33K

Released

Mar 15, 2026

Providers

1 route

The same model can have different pricing and privacy guarantees depending on who serves it. The model's headline privacy uses the strongest available route below. Discovered routes awaiting approval remain listed as not live.

ProviderPrivacyInput / 1MOutput / 1MCache read / 1MContextMax outputUptimeLatencyThroughput
VeniceVenice
Anonymous$1.20$4.00$0.24200K33K

Privacy

1 / 4

Anonymous routing

Anonymous tier

  1. Anonymous
  2. Private
  3. TEE
  4. E2EE

Your identity is hidden from the inference provider. The inference provider can see prompt content; zero retention is not guaranteed. Learn more

Features

7
Streaming
Tool calling
Reasoning
JSON/schema
Web search
Code optimized
Prompt caching

Uptime

GLM 5 Turbo — AnonRouter Models