The latest flagship model in the Qwen family. State-of-the-art results across a comprehensive suite of benchmarks — including knowledge, reasoning, coding, instruction following, human preference alignment, agent tasks, and multilingual understanding.
Modalities
In / out price
$1.20 / $6.00
per 1M
Cached price
$0.24
per 1M
Context
256K
Max output
16K
Released
–
Providers
1 routeThe same model can have different pricing and privacy guarantees depending on who serves it. The model's headline privacy uses the strongest available route below. Discovered routes awaiting approval remain listed as not live.
| Provider | Privacy | Input / 1M | Output / 1M | Cache read / 1M | Context | Max output | Uptime | Latency | Throughput |
|---|---|---|---|---|---|---|---|---|---|
| Anonymous | $1.20 | $6.00 | $0.24 | 256K | 16K | — | — | — |
Privacy
1 / 4Anonymous routing
Anonymous tier
- Anonymous
- Private
- TEE
- E2EE
Your identity is hidden from the inference provider. The inference provider can see prompt content; zero retention is not guaranteed. Learn more