Zhipu

GLM 4.7 Flash Heretic

z-ai
Chat

GLM-4.7-Flash-Heretic is an uncensored experimental variant of GLM-4.7-Flash, optimized for creative freedom and unfiltered dialogue with fast inference speed.

Modalities

In / out price

$0.07 / $0.4

per 1M

Context

200K

Max output

24K

Released

Feb 4, 2026

Updated

Jul 17, 2026

Privacy

2 / 3

Private routing

Private tier

  1. Anonymous
  2. Private
  3. E2EE
  • Prompt and response content is not retained after the request.
  • Request content is visible to the inference runtime while it is being processed.

Moderation

Uncensored

Listed as uncensored. The gateway does not add a content filter.

Data retention

Zero retention

The provider reports that request content is discarded after inference.

Features

6
Streaming
Tool calling
Reasoning
JSON/schema
Web search
Prompt caching

Routing

2

Anonrouter hosted

Routed through AnonRouter's gateway with metadata-only logging.

Provider direct

Requests egress directly to the provider runtime.

USD per 1M tokens from the Venice models API.

Providers & sources

1 route

Uptime

30d