Skip to main content

Qwen Flash

Alibaba Cloud 🧠 Language Model

Ultra-fast model with 1M context. Best latency and cost efficiency for simple tasks.

Specifications

Output modalitiestext
Capabilitiestools, reasoning
Relative speed●●●●● Blazing
Context window1,000,000 tokens
Max output8,192 tokens
Input price$0.0385 / 1M
Output price$0.378 / 1M

Pricing

$0.0385 / 1M input tokens  ·  $0.378 / 1M output tokens

Platform credit pricing — live rates on the Models page.