Qwen
Qwen3.7 Flash
Qwen3.7 Flash is a vision-language reasoning model from Alibaba.
At a glance
- Input / 1M
- $0.03
- Output / 1M
- $0.13
- Context
- 1M
- Max output
- 66K
- Intelligence
- Not available
- Released
- Jul 2026
Pricing
Qwen3.7 Flash API pricing
Per-token rates, plus what common tasks actually cost.
| Input tokensPer 1M tokens | $0.03 |
|---|---|
| Output tokensPer 1M tokens | $0.13 |
| Cached input (read)Per 1M tokens | $0.01 |
| Cache writePer 1M tokens | $0.04 |
| Long-context pricingPrompts above 32K tokens | $0.10 in / $0.40 out |
Chat message
2K in, 500 out
$0.13
per 1,000 requests
Coding task
30K in, 4K out
$1.42
per 1,000 requests
Long document summary
150K in, 2K out
$4.76
per 1,000 requests
Specs
Context window and capabilities
What it can read, what it can write, and which API features it supports.
- Context window
- 1M tokens
- Max output
- 66K tokens
- Input types
- Image, Text, Video
- Output types
- Text
- Reasoning effort
- Not available
- Release date
- July 27, 2026
- Tool calling
- Structured outputs
- JSON mode
- Reasoning
- Temperature
- Stop sequences
- Deterministic seed
- Verbosity control
Compare
Qwen3.7 Flash vs other models
See how Qwen3.7 Flash stacks up head to head on price, benchmarks, and features.
Qwen3.7 Flash vs DeepSeek V4.1 Flash
Qwen vs DeepSeek
See comparisonQwen3.7 Flash vs GLM 5.3 Flash
Qwen vs Z.ai
See comparisonQwen3.7 Flash vs MiMo-V2.6-Flash
Qwen vs Xiaomi
See comparisonQwen3.7 Flash vs GPT-6 Luna
Qwen vs OpenAI
See comparisonQwen3.7 Flash vs Hy3
Qwen vs Tencent
See comparisonQwen3.7 Flash vs Muse Spark 1.3 Contributor
Qwen vs Meta
See comparisonAlternatives
Qwen3.7 Flash alternatives
Similar models worth considering, and where each one has the edge.
Z.ai
GLM 5.3 Flash
DeepSeek
DeepSeek V4.1 Flash
OpenAI
GPT-6 Luna
Want to learn about other models? Browse all models
FAQ
Frequently asked questions
How much does Qwen3.7 Flash cost?
Qwen3.7 Flash costs $0.03 per 1M input tokens and $0.13 per 1M output tokens through the API. A typical chat message costs about $0.13 per 1,000 messages.
What is the context window of Qwen3.7 Flash?
Qwen3.7 Flash accepts up to 1M tokens of input and can write up to 66K tokens in a single response.
What is the best alternative to Qwen3.7 Flash?
GLM 5.3 Flash from Z.ai is a strong alternative. It offers: larger 1.05m context, newer release.
Can I use Qwen3.7 Flash alongside other models?
Yes. Shortcut Chat sends one prompt to Qwen3.7 Flash and any other models you pick at the same time, so you can compare answers side by side.
Try Qwen3.7 Flash next to every other model.
Use Qwen3.7 Flash alongside all the other models. Use one at a time, or multiple together. Get the best answer from every AI.