Zhipu

GLM 5.3 Flash

z-ai
Chat

GLM 5.3 Flash running in a Trusted Execution Environment (TEE). A fast, low-cost multimodal model with 1M context. Hardware attestation evidence is available for independent verification of enclave identity and configuration.

Modalities

In / out price

$0.08 / $0.27

per 1M

Cached price

$0.016

per 1M

Context

1M

Max output

33K

Released

Sep 6, 2026

Providers

4 routes

The same model can have different pricing and privacy guarantees depending on who serves it. The model's headline privacy uses the strongest available route below. Discovered routes awaiting approval remain listed as not live.

ProviderPrivacyInput / 1MOutput / 1MCache read / 1MContextMax outputUptimeLatencyThroughput
VeniceVeniceNot live
E2EE$0.08$0.27$0.0161M33K
VeniceVeniceNot live
Private$0.15$0.50$0.031M131K
DeepInfraDeepInfraNot live
Private$0.15$0.50$0.031M16K
Phala AINot live
Private$0.15$0.50$0.031M131K

Privacy

4 / 4

End-to-end encrypted routing

E2EE tier

  1. Anonymous
  2. Private
  3. TEE
  4. E2EE

Your client encrypts the prompt before it leaves your environment. E2EE includes TEE isolation; only the verified enclave can decrypt the request. Learn more

Features

7
Streaming
Tool calling
Reasoning
Vision
Web search
Code optimized
Prompt caching

Uptime

GLM 5.3 Flash — AnonRouter Models