Mercury 2 is a diffusion-based reasoning LLM from Inception, delivering over 1,000 tokens per second — 5x faster than leading speed-optimized models — with strong reasoning, tool use, and structured output capabilities.
Beta
Modalities
In / out price
$0.3125 / $0.9375
per 1M
Context
128K
Max output
50K
Released
Feb 20, 2026
Updated
Jul 17, 2026
Privacy
1 / 3Anonymous routing
Anonymous tier
- Anonymous
- Private
- E2EE
- Your identity is hidden from the inference provider.
- The inference provider can see prompt content; zero retention is not guaranteed.
Moderation
Not specified
No model-level moderation classification is recorded in this catalog.
Data retention
Not guaranteed
AnonRouter stores no payloads. Provider zero-retention is not guaranteed.
Features
6Streaming
Tool calling
Reasoning
JSON/schema
Web search
Prompt caching
Routing
2Anonrouter hosted
Routed through AnonRouter's gateway with metadata-only logging.
Provider direct
Requests egress directly to the provider runtime.
USD per 1M tokens from the Venice models API.