TAAFT
Free mode
100% free
Freemium
Free Trial
Prompts Deals

Gemini 3.5 Flash Lite

By Google
Model family: Gemini
Gemini 3.5 Flash Lite accepts text, images, audio, and video with a 1M token context window and 64K token output, running at 350 output tokens per second, the fastest in the 3.5 series. It significantly outperforms Gemini 3.1 Flash Lite on agentic and coding tasks such as Terminal Bench 2.1, long context tasks like GDM MRCR v2, and real world task execution on GDPval AA v2, in some cases outperforming the larger Gemini 3 Flash model. It is designed for high throughput production traffic including agentic search, document processing, and translation. Knowledge cutoff: March 2026.
New Multimodal Visit model
Released: July 21, 2026

Overview

Gemini 3.5 Flash Lite is Google's fastest, most cost efficient model in the Gemini 3 series. It is a natively multimodal reasoning model built on Gemini 3.1 Flash Lite, optimized for high volume, latency sensitive tasks like translation, classification, and document processing, while also supporting agentic workflows.

About Google

At Google, we think that AI can meaningfully improve people's lives and that the biggest impact will come when everyone can access it.

Industry: Technology, Information and Internet
Company Size: 190820
Location: Mountain View, CA, US
Website: ai.google
View Company Profile
Last updated: July 22, 2026
0 AIs selected
Clear selection
#
Name
Task