Models
July 21, 2026

Google introduces Gemini 3.5 Flash-Lite

Google has launched Gemini 3.5 Flash-Lite, its fastest and most affordable Flash model, optimized for high-volume, low-latency AI tasks such as translation, classification, summarization, and enterprise-scale automation.

Google DeepMind has introduced Gemini 3.5 Flash-Lite, a production-ready AI model built for high-volume, latency-sensitive workloads.

The model is optimized for tasks such as translation, document classification, summarization, and large-scale content processing while delivering lower inference costs and faster response times than larger models.

Gemini 3.5 Flash-Lite is generally available through the Gemini API, Vertex AI, and Google AI Studio, making it suitable for enterprise applications that require speed, scalability, and cost efficiency. Google positions Flash-Lite as its most economical Flash model, complementing Gemini 3.6 Flash for more advanced reasoning and agentic AI workloads.

#
Google

Read Our Content

See All Blogs
AI safety

Enterprise AI security: How GoML builds prompt injection-resistant applications

Paushigaa S

July 21, 2026
Read more
Gen AI

How Hiswai built an AI report generation software with GoML to cut research time by 80%

Deveshi Dabbawala

July 20, 2026
Read more