Google DeepMind has introduced Gemini 3.8 Flash, its latest Flash model designed for cost-effective production AI agents. Built on Gemini 3.7 Flash, the model improves software engineering, long-horizon agentic tasks and complex knowledge workflows while retaining customizable effort levels for balancing quality, latency and cost.
Gemini 3.8 Flash supports text, images, audio and video with a one-million-token context window and 64K-token output. It also supports function calling, search and computer use. Google reports strong performance across coding, finance, legal and expert-reasoning benchmarks.
The model is available through Gemini, Google AI Studio, Gemini API, Enterprise Agent Platform and Google Antigravity.




