Back to Models

Google: Gemini 3.5 Flash Lite

Googletext
8.5 / 10 Overall Rating

Developed by Google, Gemini 3.5 Flash Lite is a high-efficiency model designed for focused task execution in multi-agent workflows. It pairs a 1-million-token context window with low pricing ($0.30/1M input) and a 65,536 max output limit. While less capable in complex reasoning than GPT-4o or Claude 3.5 Sonnet, it excels at high-throughput, low-cost operations.

Context Window
1049K
Knowledge Cutoff
2025
Max Output
16K
License
Proprietary

Where it Excels (Pros)

  • 1M token context window
  • Much cheaper than GPT-4o
  • Large 65K output token limit
  • Optimized for high-volume subagent tasks

Limitations (Cons)

  • Lacks flagship-level reasoning depth
  • Text-only output generation

Benchmark Breakdown

Reasoning (MMLU)82%
Coding (HumanEval)78%
Mathematics (MATH)76%
Long-Context Retrieval95%

Pricing Matrix

Input cost / 1M tokens$0.30
Output cost / 1M tokens$2.50

Compare Google: Gemini 3.5 Flash Lite with other frontier models:

Related Skills, Tools & Automations