Back to Models

Qwen: Qwen3.7 Flash

Qwentext
9.0 / 10 Overall Rating

Qwen3.7 Flash is a high-speed multimodal vision-language model by Alibaba featuring a 1,000,000-token context window with text, image, and video input support. Optimized for visual coding, spatial reasoning, and agentic workflows, it operates at a fraction of the cost of GPT-4o and Claude 3.5 Sonnet.

Context Window
1000K
Knowledge Cutoff
2025
Max Output
16K
License
Proprietary

Where it Excels (Pros)

  • Ultra-low pricing at $0.03/1M input
  • 1M token context window
  • 64K max output token limit
  • Native text image video support

Limitations (Cons)

  • Lacks GPT-4o deep reasoning depth
  • Weaker long-form writing than Claude

Benchmark Breakdown

Reasoning (MMLU)82%
Coding (HumanEval)78%
Mathematics (MATH)76%
Long-Context Retrieval95%

Pricing Matrix

Input cost / 1M tokens$0.03
Output cost / 1M tokens$0.13

Compare Qwen: Qwen3.7 Flash with other frontier models:

Related Skills, Tools & Automations