Qwen3-VL 235B vision-language model with MoE architecture. The most powerful VL model in the Qwen series with superior visual perception, OCR, and multimodal reasoning.
Modalities
In / out price
$0.20 / $0.88
per 1M
Cached price
$0.11
per 1M
Context
262K
Max output
16K
Released
Jan 16, 2026
Providers
3 routesThe same model can have different pricing and privacy guarantees depending on who serves it. The model's headline privacy uses the strongest available route below. Discovered routes awaiting approval remain listed as not live.
| Provider | Privacy | Input / 1M | Output / 1M | Cache read / 1M | Context | Max output | Uptime | Latency | Throughput |
|---|---|---|---|---|---|---|---|---|---|
| Private | $0.21 | $1.90 | $0.10 | 128K | 16K | — | — | — | |
| Private | $0.20 | $0.88 | $0.11 | 262K | 16K | — | — | — | |
AWS BedrockNot live | Private | $0.53 | $2.66 | — | 256K | 8K | — | — | — |
Privacy
2 / 4Private routing
Private tier
- Anonymous
- Private
- TEE
- E2EE
Prompt and response content is not retained after the request. Request content is visible to the inference runtime while it is being processed. Learn more
Features
6Streaming
Tool calling
Vision
JSON/schema
Web search
Prompt caching