Qwen3.7 Max
Alibaba's flagship (API-only on DashScope, no open weights), released May 21, 2026; 1M context. Pricing corrected to $1.25/$3.75 per 1M (cached input ~$0.25).
Qwen3.7 Max strengths
- 1M context
- Strong multilingual and coding
- Deep cache discount
- Competitive frontier quality
Pricing & context
| Context window | 1M tokens (64K max output) |
| Input price /1M | $1.25 |
| Output price /1M | $3.75 |
| Modalities | text, image |
Cost guide: a typical call of about 10K input + 2K output tokens costs roughly $0.020 at list prices. Worth modelling against cheaper tiers before committing high-volume traffic.
When to choose Qwen3.7 Max
Qwen3.7 Max is best for long-context agentic workloads, multilingual apps, and Asia-Pacific deployments. If your workload is more cost-sensitive, weigh it against gpt-oss-120b (≈$0.03 input /1M) first.
Qwen3.7 Max FAQ
How much does Qwen3.7 Max cost?
Qwen3.7 Max is priced at $1.25 per 1M input tokens and $3.75 per 1M output tokens (public API list price), with a 1M tokens (64K max output) context window. A typical call of about 10K input and 2K output tokens costs roughly $0.020.
What is Qwen3.7 Max best for?
Qwen3.7 Max by Alibaba is best for long-context agentic workloads, multilingual apps, and Asia-Pacific deployments.
How does Qwen3.7 Max pricing compare to Gemini 3.6 Flash?
Qwen3.7 Max input costs $1.25 per 1M tokens versus $1.50 for Gemini 3.6 Flash, roughly 1.2x less expensive on input. Output is $3.75 vs $7.50.
Is Qwen3.7 Max multimodal?
Qwen3.7 Max supports text, image.
Other models
All models →| 01 | Claude Fable 5 | Anthropic | $10.00 | → |
| 02 | Claude Opus 5 | Anthropic | $5.00 | → |
| 03 | GPT-5.6 Sol | OpenAI | $5.00 | → |
| 04 | GPT-5.5 | OpenAI | $5.00 | → |