Qwen3.6-35B-A3B is Alibaba's latest flagship Mixture-of-Experts model, with 35B total parameters and only 3B activated per token (256 experts, 8 routed + 1 shared). Built on direct feedback from the community, Qwen3.6 prioritizes stability and real-world utility, offering developers a more intuitive, responsive, and genuinely productive coding experience.
Modalities
In / out price
$0.10 / $0.95
per 1M
Context
262K
Max output
16K
Released
–
Updated
Aug 2, 2026
Providers
1 routeThe same model can have different pricing and privacy guarantees depending on who serves it. The model's headline privacy uses the strongest available route below.
| Provider | Privacy | Input / 1M | Output / 1M | Cache read / 1M | Context | Max output | Uptime | Latency | Throughput |
|---|---|---|---|---|---|---|---|---|---|
| Private | $0.10 | $0.95 | — | 262K | 16K | — | — | — |
Privacy
2 / 3Private routing
Private tier
- Anonymous
- Private
- E2EE
Prompt and response content is not retained after the request. Request content is visible to the inference runtime while it is being processed.