Introducing Gemini 3.1 Flash ImageGoogleReleased February 26, 2026

Gemini 3.1 Flash Image

High-speed image generation and editing model engineered for conversational chats and real-time interactive design.

text-to-imageProprietary API$0.02 / imageContext: 524.288K (524,288 tokens)

Technical Specifications

Architecture Type
Low-Latency Multimodal Diffusion
Total Parameters
Undisclosed
Context Window
524.288K (524,288 tokens)
Max Output Tokens
8.192K (8,192 tokens)
Knowledge Cutoff
November 2025
Supported Modalities
vision, image-gen, text
License & Access
Google Cloud API Terms of Service

Benchmark Evaluations

Gen Eval
0.86
Latency Sec
1.2

Deep Architectural Overview

Gemini 3.1 Flash Image reduces image creation latency to approximately 1 second while upholding crisp visual details, text rendering, and style consistency.

Strengths & Considerations

Core Strengths
  • 1.2 second average generation latency
  • Inexpensive $0.02 per image pricing
  • Strong multi-turn editing accuracy
Known Limitations
  • Ultra-fine photorealism suited better on Pro Image tier

Token & API Pricing

Input Tokens (1M)$0.20
Output Tokens (1M)$0.80
Cached Input (1M)$0.0500
Pricing is verified directly against Google's developer documentation and API rate sheets.

Similar & Alternative Models

Explore other frontier models from Google and comparable reasoning engines.

Browse all models
Google

The latest evolution in the Gemini 3 family, delivering state-of-the-art software engineering (73.7% DeepSWE) and agentic enterprise knowledge workflows.

1.049M ctx$0.75 / 1M tok ($1.50 reg)
Google

Universal omnimodal generation model that accepts any combination of text, audio, image, and video to generate any combination of outputs.

1.049M ctx$0.60 / 1M tok (Universal Omni)