Qwen

Qwen 3.5 35B A3B

qwen
Chat

Qwen 3.5 35B A3B is a highly efficient MoE model with 35B total parameters and only 3B active parameters. It surpasses the larger Qwen3-235B-A22B while being 6.7x smaller, excelling at reasoning, coding, and general knowledge tasks.

Modalities

In / out price

$0.14 / $1.00

per 1M

Cached price

$0.05

per 1M

Context

262K

Max output

82K

Released

Feb 25, 2026

Providers

2 routes

The same model can have different pricing and privacy guarantees depending on who serves it. The model's headline privacy uses the strongest available route below. Discovered routes awaiting approval remain listed as not live.

ProviderPrivacyInput / 1MOutput / 1MCache read / 1MContextMax outputUptimeLatencyThroughput
VeniceVenice
Private$0.3125$1.25$0.1563256K16K
DeepInfraDeepInfra
Private$0.14$1.00$0.05262K82K

Privacy

2 / 4

Private routing

Private tier

  1. Anonymous
  2. Private
  3. TEE
  4. E2EE

Prompt and response content is not retained after the request. Request content is visible to the inference runtime while it is being processed. Learn more

Features

8
Streaming
Tool calling
Reasoning
Vision
JSON/schema
Web search
Code optimized
Prompt caching

Uptime

Qwen 3.5 35B A3B — AnonRouter Models