Back to library
AI
best practice

Switch to Gemini 3.1 Flash-Lite for Ultra-Low Latency Image Generation

Use Gemini 3.1 Flash-Lite Image for fast, cheap visual assets.

Switch to the gemini-3.1-flash-lite-image model for high-speed, cost-effective image generation and conversational editing in high-volume environments.

Google Gemini

The Scenario

A developer needs to generate thousands of dynamic product thumbnails or UI placeholders quickly and cheaply for a web application.

Before & after

The old way

Previously, developers relied on heavier Imagen models or manual editing, which could take 5–10 minutes per asset and incur higher compute costs.

With AI

Using Gemini 3.1 Flash-Lite Image, you can generate or edit assets via API in under 5 seconds, significantly reducing the cost per image.

The Prompt

I need to generate a low-cost, high-speed visual for a flash sale banner. Using the gemini-3.1-flash-lite-image model, create an image based on this description: [DESCRIBE_IMAGE_TOPIC_HERE]. Focus on high-speed rendering requirements.

The new Gemini 3.1 Flash-Lite Image (Nano Banana 2 Lite) is generally available and optimized for low-latency, high-volume image tasks.

Source

Release notes  |  Gemini API  |  Google AI for Developers
"optimized for ultra-low latency and cost-effective image generation and editing."