Qwen 3.8 Flash is the latest multimodal model in the Qwen family, pairing strong reasoning and generation with remarkable speed. It natively supports a 1M-token context window, so it can process long documents, entire codebases, and complex conversations in a single pass. It excels at coding assistance, agentic workflows, and visual understanding, accepting text, image, and video input, and its thinking mode can be turned on or off per request.
Modalities
In / out price
$0.14 / $0.49
per 1M
Cached price
$0.014
per 1M
Context
1M
Max output
131K
Released
Sep 10, 2026
Providers
0 routesThe same model can have different pricing and privacy guarantees depending on who serves it. Only routes currently callable through AnonRouter are shown.
No provider route for this model is live right now.
Privacy
0 / 4No live route
AnonRouter is not currently offering a callable provider route for this model.