Qwen

Qwen 3.8 Max

qwen
Chat

Qwen 3.8 Max is Alibaba's flagship 2.4-trillion-parameter MoE model, with major gains over Qwen 3.7 Max in software engineering and office-productivity workflows and strong long-horizon, multi-agent performance. It accepts both text and vision-language input (images and video), operates in thinking mode only, and supports a 1M-token context window.

Modalities

In / out price

$1.65 / $4.951

per 1M

Cached price

$0.206

per 1M

Context

256K

Max output

16K

Released

Jul 22, 2026

Providers

2 routes

The same model can have different pricing and privacy guarantees depending on who serves it. The model's headline privacy uses the strongest available route below. Discovered routes awaiting approval remain listed as not live.

ProviderPrivacyInput / 1MOutput / 1MCache read / 1MContextMax outputUptimeLatencyThroughput
VeniceVenice
Anonymous$2.50$7.50$0.31251M131K
DeepInfraDeepInfra
Anonymous$1.65$4.951$0.206256K16K

Privacy

1 / 4

Anonymous routing

Anonymous tier

  1. Anonymous
  2. Private
  3. TEE
  4. E2EE

Your identity is hidden from the inference provider. The inference provider can see prompt content; zero retention is not guaranteed. Learn more

Features

7
Streaming
Tool calling
Reasoning
Vision
Web search
Code optimized
Prompt caching

Uptime

Qwen 3.8 Max — AnonRouter Models