Gemini

Gemini 3.5 Flash

google
Chat

Gemini 3.5 Flash is a high speed, high value thinking model with 1M context, designed for agentic workflows, multi-turn chat, and coding assistance. It delivers near Pro level reasoning with substantially lower latency.

Modalities

In / out price

$1.55 / $9.45

per 1M

Context

1M

Max output

66K

Released

May 22, 2026

Updated

Jul 17, 2026

Privacy

1 / 3

Anonymous routing

Anonymous tier

  1. Anonymous
  2. Private
  3. E2EE
  • Your identity is hidden from the inference provider.
  • The inference provider can see prompt content; zero retention is not guaranteed.

Moderation

Not specified

No model-level moderation classification is recorded in this catalog.

Data retention

Not guaranteed

AnonRouter stores no payloads. Provider zero-retention is not guaranteed.

Features

7
Streaming
Tool calling
Reasoning
Vision
JSON/schema
Web search
Prompt caching

Routing

2

Anonrouter hosted

Routed through AnonRouter's gateway with metadata-only logging.

Provider direct

Requests egress directly to the provider runtime.

USD per 1M tokens from the Venice models API.

Providers & sources

1 route

Uptime

30d