Qwen

Qwen3 32B

qwen
Chat

Qwen3 is the latest generation of large language models in Qwen series, offering a comprehensive suite of dense and mixture-of-experts (MoE) models. Built upon extensive training, Qwen3 delivers groundbreaking advancements in reasoning, instruction-following, agent capabilities, and multilingual support

Modalities

In / out price

$0.08 / $0.28

per 1M

Context

41K

Max output

16K

Released

Updated

Sep 19, 2026

Providers

1 route

The same model can have different pricing and privacy guarantees depending on who serves it. Only routes currently callable through AnonRouter are shown.

ProviderPrivacyInput / 1MOutput / 1MCache read / 1MContextMax outputUptimeLatencyThroughput
DeepInfraDeepInfra
Private$0.08$0.2841K16K

Privacy

2 / 4

Private routing

Private tier

  1. Anonymous
  2. Private
  3. TEE
  4. E2EE

Prompt and response content is not retained after the request. Request content is visible to the inference runtime while it is being processed. Learn more

Features

3
Streaming
Tool calling
Reasoning

Uptime

Qwen3 32B — AnonRouter Models