Back to Models
Google: Gemini 3.6 Flash
Googletext9.1 / 10 Overall Rating
Gemini 3.6 Flash is Google's high-efficiency model optimized for web development, agentic workflows, and fast multimodal processing. It combines a 1-million-token context window with a large 65,536 max output limit at a fraction of GPT-4o's API costs. The model trades top-tier complex reasoning for speed, massive context volume, and cost efficiency.
Context Window
1049K
Knowledge Cutoff
2025
Max Output
16K
License
Proprietary
Where it Excels (Pros)
- 1M token context window
- 65K max output token limit
- Significantly cheaper than GPT-4o
- Native audio and video inputs
Limitations (Cons)
- Lacks Claude 3.5 Sonnet reasoning
- Lower precision on complex tasks
Benchmark Breakdown
Reasoning (MMLU)82%
Coding (HumanEval)78%
Mathematics (MATH)76%
Long-Context Retrieval95%
Pricing Matrix
Input cost / 1M tokens$0.75
Output cost / 1M tokens$3.75
Compare Google: Gemini 3.6 Flash with other frontier models: