DeepSeek V4 Flash is an efficiency-focused MoE model with 284B total parameters (13B active) and a 1M-token context window. It's tuned for fast inference and high-throughput use cases while still holding up on reasoning and coding tasks.
Modalities
In / out price
$0.09 / $0.18
per 1M
Cached price
$0.018
per 1M
Context
1M
Max output
66K
Released
–
Providers
3 routesThe same model can have different pricing and privacy guarantees depending on who serves it. The model's headline privacy uses the strongest available route below. Discovered routes awaiting approval remain listed as not live.
Privacy
4 / 4End-to-end encrypted routing
E2EE tier
- Anonymous
- Private
- TEE
- E2EE
Your client encrypts the prompt before it leaves your environment. E2EE includes TEE isolation; only the verified enclave can decrypt the request. Learn more