Qwen 2.5 7B Instruct running in a Trusted Execution Environment (TEE). A compact model with strong coding, math, and multilingual capabilities supporting 29+ languages, with hardware attestation…
Qwen
qwen/ · 28 models
Access 28 Qwen models through AnonRouter's privacy-first gateway, including Qwen 2.5 7B, Qwen 3 235B A22B Instruct 2507, and Qwen 3 235B A22B Thinking 2507. Compare pricing, context windows, and capabilities across Qwen's text, image, audio, embeddings routes — every request anonymized, with no payload logging.
Models
28
Modalities
4
Text, Image, Audio, Embeddings
From (input)
$0.0125
per 1M tokens
Max context
1M
Private routes
19 / 28
not anonymous-only
Catalog by modality
28 routesQwen models28
Built for in-depth research and handling long, complex documents. Ideal for technical work, multimodal input, and high-precision tasks.
Built for in-depth research and handling long, complex documents. Ideal for technical work, multimodal input, and high-precision tasks.
Turbo variant of Qwen3 Coder 480B, optimized for faster inference on code tasks.
Optimized for speed and efficiency.
Qwen 3.5 35B A3B is a highly efficient MoE model with 35B total parameters and only 3B active parameters. It surpasses the larger Qwen3-235B-A22B while being 6.7x smaller, excelling at reasoning,…
Qwen 3.5 is Alibaba flagship reasoning model featuring a 397B parameter Mixture-of-Experts architecture with 17B active parameters. It excels at complex reasoning, coding, and general knowledge tasks.
A 9B dense model with 262K native context window (extendable to 1M). Features Gated DeltaNet hybrid attention architecture for efficient long-context processing. Supports 201 languages,…
The Qwen 3.6 27B native vision-language dense model builds upon the 3.5-27B version, with key improvements in agentic coding capabilities and enhanced STEM reasoning and inference skills. In the…
Qwen 3.6 27B FP8 running in a Trusted Execution Environment (TEE). Hardware attestation evidence is available for independent verification of enclave identity and configuration.
Qwen 3.6 35B A3B FP8 running in a Trusted Execution Environment (TEE). A fast mixture-of-experts model with ~3B active parameters per token. Hardware attestation evidence is available for independent…
Qwen 3.6 Plus Uncensored is Alibaba's latest flagship reasoning model with exceptional performance across coding, reasoning, and general knowledge tasks. Features mixed reasoning, function calling,…
Qwen 3.7 Max is the largest model in the Qwen 3.7 series, with deep thinking, function calling, prompt caching, and multimodal input support for images and video. It excels at programming, office and…
Qwen 3.7 Plus is Alibaba's latest flagship reasoning model with exceptional performance across coding, reasoning, and general knowledge tasks. Features mixed reasoning, function calling, and…
Qwen3 30B A3B running in a Trusted Execution Environment (TEE). A MoE model with 30.5B total parameters and 3.3B activated per inference, supporting ultra-long 256K context, with hardware attestation…
Qwen3-VL 235B vision-language model with MoE architecture. The most powerful VL model in the Qwen series with superior visual perception, OCR, and multimodal reasoning.
Qwen3 VL 30B A3B running in a Trusted Execution Environment (TEE). A multimodal model unifying text generation with visual understanding for images and videos, with hardware attestation evidence…
Qwen3.6 35B A3B Uncensored running in a Trusted Execution Environment (TEE). An uncensored variant of Alibaba's Qwen3.6 MoE model with 35B total parameters and ~3B active, supporting 262K context and…
Text-to-speech model with 9 voices served through Venice.
Text-to-speech model with 9 voices served through Venice.
Image inpainting and editing model served through Venice.
Image generation model served through Venice.
Image generation model served through Venice.
Image inpainting and editing model served through Venice.
Image generation model served through Venice.
Image inpainting and editing model served through Venice.
Embedding model for semantic search and retrieval, served through Venice.
Embedding model for semantic search and retrieval, served through Venice.