Qwen

Qwen

qwen/ · 32 models

Access 32 Qwen models through AnonRouter's privacy-first gateway, including Qwen 2.5 7B, Qwen 3 235B A22B Instruct 2507, and Qwen 3 235B A22B Thinking 2507. Compare pricing, context windows, and capabilities across Qwen's text, image, audio, embeddings routes — every request anonymized, with no payload logging.

Models

32

Modalities

4

Text, Image, Audio, Embeddings

From (input)

$0.0125

per 1M tokens

Max context

1M

Private routes

23 / 32

not anonymous-only

Catalog by modality

32 routes
Text22Image6Audio2Embeddings2

Qwen models32

Qwen
Qwen 2.5 7B
qwen/qwen-2.5-7b
E2EE

Qwen 2.5 7B Instruct running in a Trusted Execution Environment (TEE). A compact model with strong coding, math, and multilingual capabilities supporting 29+ languages, with hardware attestation evidence available for independent verification.

E2EE|32K context|$0.05/M input|$0.13/M output
Qwen
Qwen 3 235B A22B Instruct 2507
qwen/qwen-3-235b-a22b-instruct-2507
Private

Built for in-depth research and handling long, complex documents. Ideal for technical work, multimodal input, and high-precision tasks.

Private|128K context|$0.15/M input|$0.75/M output
Qwen
Qwen 3 235B A22B Thinking 2507
qwen/qwen-3-235b-a22b-thinking-2507
Private

Built for in-depth research and handling long, complex documents. Ideal for technical work, multimodal input, and high-precision tasks.

Private|128K context|$0.45/M input|$3.50/M output
Qwen
Qwen 3 Coder 480B Turbo
qwen/qwen-3-coder-480b-turbo
Private

Turbo variant of Qwen3 Coder 480B, optimized for faster inference on code tasks.

Private|256K context|$0.35/M input|$1.50/M output
Qwen
Qwen 3 Next 80b
qwen/qwen-3-next-80b
Private

Optimized for speed and efficiency.

Private|256K context|$0.35/M input|$1.90/M output
Qwen
Qwen 3.5 35B A3B
qwen/qwen-3.5-35b-a3b
Private

Qwen 3.5 35B A3B is a highly efficient MoE model with 35B total parameters and only 3B active parameters. It surpasses the larger Qwen3-235B-A22B while being 6.7x smaller, excelling at reasoning, coding, and general knowledge tasks.

Private|256K context|$0.3125/M input|$1.25/M output
Qwen
Qwen 3.5 397B
qwen/qwen-3.5-397b
Anonymous

Qwen 3.5 is Alibaba flagship reasoning model featuring a 397B parameter Mixture-of-Experts architecture with 17B active parameters. It excels at complex reasoning, coding, and general knowledge tasks.

Anonymous|128K context|$0.75/M input|$4.50/M output
Qwen
Qwen 3.5 9B
qwen/qwen-3.5-9b
Private

A 9B dense model with 262K native context window (extendable to 1M). Features Gated DeltaNet hybrid attention architecture for efficient long-context processing. Supports 201 languages, thinking/reasoning mode, and function calling.

Private|256K context|$0.10/M input|$0.15/M output
Qwen
Qwen 3.6 27B
qwen/qwen-3.6-27b
Private

The Qwen 3.6 27B native vision-language dense model builds upon the 3.5-27B version, with key improvements in agentic coding capabilities and enhanced STEM reasoning and inference skills. In the vision modality, it demonstrates significant advances in spatial intelligence, object localization, and detection, while video understanding, document OCR, and visual agent capabilities continue to improve steadily.

Private|256K context|$0.325/M input|$3.25/M output
Qwen
Qwen 3.6 27B FP8
qwen/qwen-3.6-27b-fp8
E2EE

Qwen 3.6 27B FP8 running in a Trusted Execution Environment (TEE). Hardware attestation evidence is available for independent verification of enclave identity and configuration.

E2EE|256K context|$0.346/M input|$3.46/M output
Qwen
Qwen 3.6 35B A3B FP8
qwen/qwen-3.6-35b-a3b-fp8
E2EE

Qwen 3.6 35B A3B FP8 running in a Trusted Execution Environment (TEE). A fast mixture-of-experts model with ~3B active parameters per token. Hardware attestation evidence is available for independent verification of enclave identity and configuration.

E2EE|32K context|$0.182/M input|$1.18/M output
Qwen
Qwen 3.6 Plus Uncensored
qwen/qwen-3.6-plus-uncensored
Anonymous

Qwen 3.6 Plus Uncensored is Alibaba's latest flagship reasoning model with exceptional performance across coding, reasoning, and general knowledge tasks. Features mixed reasoning, function calling, and multimodal input support.

Anonymous|1M context|$0.625/M input|$3.75/M output
Qwen
Qwen 3.7 Max
qwen/qwen-3.7-max
Anonymous

Qwen 3.7 Max is the largest model in the Qwen 3.7 series, with deep thinking, function calling, prompt caching, and multimodal input support for images and video. It excels at programming, office and productivity tasks, and long-running autonomous agent workflows.

Anonymous|1M context|$2.70/M input|$8.05/M output
Qwen
Qwen 3.7 Plus
qwen/qwen-3.7-plus
Anonymous

Qwen 3.7 Plus is Alibaba's latest flagship reasoning model with exceptional performance across coding, reasoning, and general knowledge tasks. Features mixed reasoning, function calling, and multimodal input support.

Anonymous|1M context|$0.50/M input|$2.00/M output
Qwen
Qwen3 30B A3B
qwen/qwen3-30b-a3b
E2EE

Qwen3 30B A3B running in a Trusted Execution Environment (TEE). A MoE model with 30.5B total parameters and 3.3B activated per inference, supporting ultra-long 256K context, with hardware attestation evidence available for independent verification.

E2EE|256K context|$0.19/M input|$0.69/M output
Qwen
Qwen3 VL 235B
qwen/qwen3-vl-235b
Private

Qwen3-VL 235B vision-language model with MoE architecture. The most powerful VL model in the Qwen series with superior visual perception, OCR, and multimodal reasoning.

Private|128K context|$0.21/M input|$1.90/M output
Qwen
Qwen3 VL 30B A3B
qwen/qwen3-vl-30b-a3b
E2EE

Qwen3 VL 30B A3B running in a Trusted Execution Environment (TEE). A multimodal model unifying text generation with visual understanding for images and videos, with hardware attestation evidence available for independent verification.

E2EE|128K context|$0.25/M input|$0.90/M output
Qwen
Qwen3.6 35B A3B Uncensored
qwen/qwen3.6-35b-a3b-uncensored
E2EE

Qwen3.6 35B A3B Uncensored running in a Trusted Execution Environment (TEE). An uncensored variant of Alibaba's Qwen3.6 MoE model with 35B total parameters and ~3B active, supporting 262K context and multimodal input across text, images, and video, with hardware attestation evidence available for independent verification.

E2EE|128K context|$0.38/M input|$1.88/M output
Qwen
Qwen3 32B
qwen/qwen3-32b
E2EE

Qwen3 32B served through Amazon Bedrock with support for AWS zero-data-retention mode.

E2EE|32K context|$0.15/M input|$0.60/M output
Qwen
Qwen3 Coder 30B A3B Instruct
qwen/qwen3-coder-30b-a3b-instruct
Private

Qwen3 Coder 30B A3B Instruct served through Amazon Bedrock with support for AWS zero-data-retention mode.

Private|256K context|$0.15/M input|$0.60/M output
Qwen
Qwen3 Coder Next
qwen/qwen3-coder-next
Private

Qwen3 Coder Next served through Amazon Bedrock with support for AWS zero-data-retention mode.

Private|256K context|$0.50/M input|$1.20/M output
Qwen
Qwen3.6 35B A3B
qwen/qwen-3.6-35b-a3b
Private

Qwen3.6 35B A3B served on DeepInfra serverless inference.

Private|262K context|$0.10/M input|$0.95/M output
Qwen
Qwen 3 TTS 0.6B
qwen/qwen-3-tts-0.6b
Private

Text-to-speech model with 9 voices served through Venice.

Private|Unknown context|$87.50/1M chars|
Qwen
Qwen 3 TTS 1.7B
qwen/qwen-3-tts-1.7b
Private

Text-to-speech model with 9 voices served through Venice.

Private|Unknown context|$112.50/1M chars|
Qwen
Qwen Edit Uncensored
qwen/qwen-edit-uncensored
Private

Image inpainting and editing model served through Venice.

Private|Unknown context||$0.04/edit
Qwen
Qwen Image
qwen/qwen-image
Anonymous

Image generation model served through Venice.

Anonymous|Unknown context||$0.03/image
Qwen
Qwen Image 2
qwen/qwen-image-2
Anonymous

Image generation model served through Venice.

Anonymous|Unknown context||$0.05/image
Qwen
Qwen Image 2
qwen/qwen-image-2-edit
Anonymous

Image inpainting and editing model served through Venice.

Anonymous|Unknown context||$0.05/edit
Qwen
Qwen Image 2 Pro
qwen/qwen-image-2-pro
Anonymous

Image generation model served through Venice.

Anonymous|Unknown context||$0.10/image
Qwen
Qwen Image 2 Pro
qwen/qwen-image-2-pro-edit
Anonymous

Image inpainting and editing model served through Venice.

Anonymous|Unknown context||$0.10/edit
Qwen
Qwen3 Embedding 0.6B
qwen/qwen3-embedding-0.6b
Private

Embedding model for semantic search and retrieval, served through Venice.

Private|33K context|$0.0125/M input|$0.0125/M output
Qwen
Qwen3 Embedding 8B
qwen/qwen3-embedding-8b
Private

Embedding model for semantic search and retrieval, served through Venice.

Private|33K context|$0.0125/M input|$0.0125/M output
Qwen models — AnonRouter