← All model drops
Live drop
3.5 Flash-Lite
Google DeepMindDropped Jul 21, 2026
350 tokens/sec at $0.30/$2.50 for high-volume sub-agent and document pipelines.
About this release
Live drop
Google DeepMindDropped Jul 21, 2026
350 tokens/sec at $0.30/$2.50 for high-volume sub-agent and document pipelines.
About this release
Gemini 3.5 Flash-Lite is the fastest, lowest-cost 3.5-class model, optimized for high-throughput agentic search, data extraction, and repetitive sub-tasks. It outperforms prior Lite generations on agentic and coding benchmarks (e.g., Terminal-Bench, OSWorld) while running at 350 output tokens per second and supporting computer-use tooling.
Read the official Google DeepMind launch pageFirst builds incoming — this lane fills in automatically as people start building with 3.5 Flash-Lite. Know one? Submit the URL.