Model
Qwen Flash
Efficient Qwen model for fast chat, extraction, and high-volume workloads
2 providersreleased knowledge cutoff 2024-04closed weights
$0.05
best input /M · LLM Gateway
$0.4
best output /M
1M
context window
33K
max output
Compare providers (sorted by listed input price)
| Provider | Input /M | Output /M | Cache read /M | Cache write /M | Context | Capabilities | Status | Updated |
|---|---|---|---|---|---|---|---|---|
| LLM Gateway alibaba/qwen-flash | $0.05 | $0.4 | $0.01 | $0.063 | 1M | no reasoningtoolsno structuredno vision | Jul 28, 2025 | |
| Alibaba qwen-flash | $0.05 | $0.4 | — | — | 1M | reasoningtoolsno structuredno vision | Jul 28, 2025 |
Price badge
https://models.sutraworks.ai/badge/alibaba/qwen-flash.svgLive SVG, regenerated on every hourly sync — paste it into a README to always show the current cheapest listed price. Updates when prices change; no build step needed on your side.
Prices are per million tokens (USD). “—” means the provider does not publicly list a price for this model. Data from models.dev; verify with the provider before purchasing.