Models
September 2, 2026

Google introduces Gemini 3.8 Flash for coding and agentic AI workloads

Google has introduced Gemini 3.8 Flash, improving software engineering, agentic workflows and knowledge work while offering customizable reasoning effort, multimodal inputs, a one-million-token context window and 64K-token output.

Google DeepMind has introduced Gemini 3.8 Flash, its latest Flash model designed for cost-effective production AI agents. Built on Gemini 3.7 Flash, the model improves software engineering, long-horizon agentic tasks and complex knowledge workflows while retaining customizable effort levels for balancing quality, latency and cost.

Gemini 3.8 Flash supports text, images, audio and video with a one-million-token context window and 64K-token output. It also supports function calling, search and computer use. Google reports strong performance across coding, finance, legal and expert-reasoning benchmarks.

The model is available through Gemini, Google AI Studio, Gemini API, Enterprise Agent Platform and Google Antigravity.

#
Google

Read Our Content

See All Blogs
LLM Models

GPT 6 LLM testing: What the benchmarks mean for enterprise delivery

Sarankumar S

September 8, 2026
Read more
LLM Models

LLM testing of Gemini 3.8 Flash: Scored on the AI Matic Bench

Sarankumar S

September 4, 2026
Read more