Back to Models
Google: Gemini 3.7 Flash
Googletext9.2 / 10 Overall Rating
Gemini 3.7 Flash is a multimodal model from Google built for fast agentic execution, complex coding, and multi-step reasoning. It features a 1,048,576 token input context window, a 65,536 token max output window, and native support for audio, video, image, and text inputs. Priced at $0.38 per million input tokens, it offers low-cost, high-throughput execution compared to GPT-4o and Claude 3.5 Sonnet.
Context Window
1049K
Knowledge Cutoff
2025
Max Output
16K
License
Proprietary
Where it Excels (Pros)
- 1M token input context window
- 65K max output token limit
- Significantly cheaper than GPT-4o
- Native audio and video inputs
Limitations (Cons)
- Reasoning depth trails frontier benchmarks
- Inconsistent recall at maximum context
Benchmark Breakdown
Reasoning (MMLU)88%
Coding (HumanEval)86%
Mathematics (MATH)84%
Long-Context Retrieval95%
Pricing Matrix
Input cost / 1M tokens$0.38
Output cost / 1M tokens$1.88
Compare Google: Gemini 3.7 Flash with other frontier models: