Back to Models

Google: Gemini 3.6 Flash

Googletext
9.1 / 10 Overall Rating

Gemini 3.6 Flash is Google's high-efficiency model optimized for web development, agentic workflows, and fast multimodal processing. It combines a 1-million-token context window with a large 65,536 max output limit at a fraction of GPT-4o's API costs. The model trades top-tier complex reasoning for speed, massive context volume, and cost efficiency.

Context Window
1049K
Knowledge Cutoff
2025
Max Output
16K
License
Proprietary

Where it Excels (Pros)

  • 1M token context window
  • 65K max output token limit
  • Significantly cheaper than GPT-4o
  • Native audio and video inputs

Limitations (Cons)

  • Lacks Claude 3.5 Sonnet reasoning
  • Lower precision on complex tasks

Benchmark Breakdown

Reasoning (MMLU)82%
Coding (HumanEval)78%
Mathematics (MATH)76%
Long-Context Retrieval95%

Pricing Matrix

Input cost / 1M tokens$0.75
Output cost / 1M tokens$3.75

Compare Google: Gemini 3.6 Flash with other frontier models:

Related Skills, Tools & Automations