콘텐츠로 이동

Models#

Skip to main content

/

  • English
  • Deutsch
  • Español – América Latina
  • Français
  • Indonesia
  • Italiano
  • Polski
  • Português – Brasil
  • Shqip
  • Tiếng Việt
  • Türkçe
  • Русский
  • עברית
  • العربيّة
  • فارسی
  • हिंदी
  • বাংলা
  • ภาษาไทย
  • 中文 – 简体
  • 中文 – 繁體
  • 日本語
  • 한국어

Get API key Cookbook Community Sign in

Docs API reference

The Interactions API is now generally available. We recommend using this API for access to all the latest features and models.

Send feedback

Models#

This guide introduces all the models available through the Gemini API.


Gemini 3#

Stable#

spark Gemini 3.6 Flash Our latest model that balances speed with intelligence to deliver strong performance in agentic and multimodal tasks. Stable spark Gemini 3.5 Flash Most intelligent model for sustained frontier performance on agentic and coding tasks. Stable bolt Gemini 3.5 Flash-Lite Our fastest, most cost-effective 3.5 model for high-throughput execution. Stable bolt Gemini 3.1 Flash-Lite Frontier-class performance rivaling larger models at a fraction of the cost. Stable 🍌🍌 Nano Banana 2 Powerful, high-efficiency image generation and editing, optimized for speed and high-volume use cases. Stable 🍌 Nano Banana 2 Lite Ultra-low latency and cost-effective image generation and editing, designed for high-volume interactive use cases. Stable 🍌 Nano Banana Pro State-of-the-art image generation and editing models for highly contextual native image creation. Stable

Preview#

auto_awesome Gemini 3.1 Pro Advanced intelligence, complex problem-solving skills, and powerful agentic and vibe coding capabilities. Preview spark Gemini 3 Flash Frontier-class performance rivaling larger models at a fraction of the cost. Preview translate Gemini 3.5 Live Translate Low-latency, real-time speech to speech translation model that supports 70+ languages. New Preview settings_voice Gemini 3.1 Flash Live High-quality, low-latency Live API model for real-time dialogue and voice-first AI applications. New Preview graphic_eq Gemini 3.1 Flash TTS Powerful, low-latency speech generation. New Preview movie_filter Gemini Omni Flash Fast, conversational video generation and editing. Turn text and images into video, and refine results through natural language. New Preview

All Gemini 3 models#

Model Endpoint
Gemini 3.6 Flash
gemini-3.6-flash

Gemini 3.5 Flash |

gemini-3.5-flash

Gemini 3.5 Flash-Lite |

gemini-3.5-flash-lite

Gemini 3.1 Flash-Lite |

gemini-3.1-flash-lite

Nano Banana 2 |

gemini-3.1-flash-image

Nano Banana 2 Lite |

gemini-3.1-flash-lite-image

Nano Banana Pro |

gemini-3-pro-image

Gemini 3.1 Pro |

gemini-3.1-pro-preview

Gemini 3 Flash |

gemini-3-flash-preview

Gemini 3.5 Live Translate |

gemini-3.5-live-translate-preview

Gemini 3.1 Flash Live |

gemini-3.1-flash-live-preview

Gemini 3.1 Flash TTS |

gemini-3.1-flash-tts-preview

Gemini Omni Flash |

gemini-omni-flash

Gemini 2.5 Flash#

Model Description Endpoint
Gemini 2.5 Flash Our best price-performance model for low-latency, high-volume tasks that require reasoning.
gemini-2.5-flash

Nano Banana | State-of-the-art native image generation and editing designed for fast, creative workflows. |

gemini-2.5-flash-image

Gemini 2.5 Flash Live | Optimized for real-time conversational agents with sub-second native audio streaming. |

gemini-2.5-flash-native-audio-preview-12-2025

Gemini 2.5 Flash TTS | Controllable text-to-speech audio generation with fine control over style and pacing. |

gemini-2.5-flash-preview-tts

Gemini 2.5 Flash-Lite#

Model Description Endpoint
Gemini 2.5 Flash-Lite The fastest and most budget-friendly multimodal model in the 2.5 family.
gemini-2.5-flash-lite

Gemini 2.5 Pro#

Model Description Endpoint
Gemini 2.5 Pro Our most advanced model for complex tasks, featuring deep reasoning and coding capabilities in the 2.5 family.
gemini-2.5-pro

Gemini 2.5 Pro TTS | High-fidelity speech synthesis optimized for quality in structured workflows like podcasts and audiobooks. |

gemini-2.5-pro-preview-tts

Audio models#

This section contains all audio models, including ones that may already be listed in other sections

Model Description Endpoint
Gemini 3.1 Flash Live Our high-quality, low-latency audio-to-audio (A2A) model designed for real-time dialogue and voice-first AI applications.
gemini-3.1-flash-live-preview

Gemini 3.1 Flash TTS | Powerful, low-latency speech generation, with natural outputs, steerable prompts, and new expressive audio tags for precise narration control. |

gemini-3.1-flash-tts-preview

Gemini 2.5 Flash Live | Our flagship Live API model for low-latency, bidirectional voice and video agents with native audio reasoning. |

gemini-2.5-flash-native-audio-preview-12-2025

Gemini 2.5 Flash TTS | Fast and controllable text-to-speech for low-latency, cost-efficient applications and real-time assistants. |

gemini-2.5-flash-preview-tts

Gemini 2.5 Pro TTS | High-fidelity speech synthesis optimized for quality in str