Back to Models

Google: Gemini 3.7 Flash

Googletext
9.2 / 10 Overall Rating

Gemini 3.7 Flash is a multimodal model from Google built for fast agentic execution, complex coding, and multi-step reasoning. It features a 1,048,576 token input context window, a 65,536 token max output window, and native support for audio, video, image, and text inputs. Priced at $0.38 per million input tokens, it offers low-cost, high-throughput execution compared to GPT-4o and Claude 3.5 Sonnet.

Context Window
1049K
Knowledge Cutoff
2025
Max Output
16K
License
Proprietary

Where it Excels (Pros)

  • 1M token input context window
  • 65K max output token limit
  • Significantly cheaper than GPT-4o
  • Native audio and video inputs

Limitations (Cons)

  • Reasoning depth trails frontier benchmarks
  • Inconsistent recall at maximum context

Benchmark Breakdown

Reasoning (MMLU)88%
Coding (HumanEval)86%
Mathematics (MATH)84%
Long-Context Retrieval95%

Pricing Matrix

Input cost / 1M tokens$0.38
Output cost / 1M tokens$1.88

Compare Google: Gemini 3.7 Flash with other frontier models:

Related Skills, Tools & Automations