Inception

Mercury 2

inception
Chat

Mercury 2 is a diffusion-based reasoning LLM from Inception, delivering over 1,000 tokens per second — 5x faster than leading speed-optimized models — with strong reasoning, tool use, and structured output capabilities.

Beta

Modalities

In / out price

$0.3125 / $0.9375

per 1M

Context

128K

Max output

50K

Released

Feb 20, 2026

Updated

Jul 17, 2026

Privacy

1 / 3

Anonymous routing

Anonymous tier

  1. Anonymous
  2. Private
  3. E2EE
  • Your identity is hidden from the inference provider.
  • The inference provider can see prompt content; zero retention is not guaranteed.

Moderation

Not specified

No model-level moderation classification is recorded in this catalog.

Data retention

Not guaranteed

AnonRouter stores no payloads. Provider zero-retention is not guaranteed.

Features

6
Streaming
Tool calling
Reasoning
JSON/schema
Web search
Prompt caching

Routing

2

Anonrouter hosted

Routed through AnonRouter's gateway with metadata-only logging.

Provider direct

Requests egress directly to the provider runtime.

USD per 1M tokens from the Venice models API.

Providers & sources

1 route

Uptime

30d