Back to Models
Google: Gemini 3.5 Flash Lite
Googletext8.5 / 10 Overall Rating
Developed by Google, Gemini 3.5 Flash Lite is a high-efficiency model designed for focused task execution in multi-agent workflows. It pairs a 1-million-token context window with low pricing ($0.30/1M input) and a 65,536 max output limit. While less capable in complex reasoning than GPT-4o or Claude 3.5 Sonnet, it excels at high-throughput, low-cost operations.
Context Window
1049K
Knowledge Cutoff
2025
Max Output
16K
License
Proprietary
Where it Excels (Pros)
- 1M token context window
- Much cheaper than GPT-4o
- Large 65K output token limit
- Optimized for high-volume subagent tasks
Limitations (Cons)
- Lacks flagship-level reasoning depth
- Text-only output generation
Benchmark Breakdown
Reasoning (MMLU)82%
Coding (HumanEval)78%
Mathematics (MATH)76%
Long-Context Retrieval95%
Pricing Matrix
Input cost / 1M tokens$0.30
Output cost / 1M tokens$2.50
Compare Google: Gemini 3.5 Flash Lite with other frontier models: