Meet Qwen3-VL — the most powerful vision-language model in the Qwen series to date. This generation delivers comprehensive upgrades across the board: superior text understanding & generation, deeper visual perception & reasoning, extended context length, enhanced spatial and video dynamics comprehension, and stronger agent interaction capabilities.
Modalities
In / out price
$0.15 / $0.60
per 1M
Context
262K
Max output
16K
Released
–
Updated
Sep 19, 2026
Providers
1 routeThe same model can have different pricing and privacy guarantees depending on who serves it. Only routes currently callable through AnonRouter are shown.
| Provider | Privacy | Input / 1M | Output / 1M | Cache read / 1M | Context | Max output | Uptime | Latency | Throughput |
|---|---|---|---|---|---|---|---|---|---|
| Private | $0.15 | $0.60 | — | 262K | 16K | — | — | — |
Privacy
2 / 4Private routing
Private tier
- Anonymous
- Private
- TEE
- E2EE
Prompt and response content is not retained after the request. Request content is visible to the inference runtime while it is being processed. Learn more