Back to Models
Qwen: Qwen3.7 Flash
Qwentext9.0 / 10 Overall Rating
Qwen3.7 Flash is a high-speed multimodal vision-language model by Alibaba featuring a 1,000,000-token context window with text, image, and video input support. Optimized for visual coding, spatial reasoning, and agentic workflows, it operates at a fraction of the cost of GPT-4o and Claude 3.5 Sonnet.
Context Window
1000K
Knowledge Cutoff
2025
Max Output
16K
License
Proprietary
Where it Excels (Pros)
- Ultra-low pricing at $0.03/1M input
- 1M token context window
- 64K max output token limit
- Native text image video support
Limitations (Cons)
- Lacks GPT-4o deep reasoning depth
- Weaker long-form writing than Claude
Benchmark Breakdown
Reasoning (MMLU)82%
Coding (HumanEval)78%
Mathematics (MATH)76%
Long-Context Retrieval95%
Pricing Matrix
Input cost / 1M tokens$0.03
Output cost / 1M tokens$0.13
Compare Qwen: Qwen3.7 Flash with other frontier models: