Models#
Skip to main content
/
- English
- Deutsch
- Español – América Latina
- Français
- Indonesia
- Italiano
- Polski
- Português – Brasil
- Shqip
- Tiếng Việt
- Türkçe
- Русский
- עברית
- العربيّة
- فارسی
- हिंदी
- বাংলা
- ภาษาไทย
- 中文 – 简体
- 中文 – 繁體
- 日本語
- 한국어
Get API key Cookbook Community Sign in
- Get API key
- Cookbook
-
Get started
- Get Started
- API keys
- Pricing
- Coding agent setup
-
Models
- Latest Gemini models
- Nano Banana
- Veo
- Gemini Omni Flash
- Lyria 3
- Lyria RealTime
- Imagen
- Text-to-speech
- Live
- Live Translate
- Embeddings
-
Robotics
-
Core capabilities
-
Image
-
Video
-
Speech and audio
-
Thinking
- Function calling
- Long context
-
Agents
- Quickstart
- Antigravity Agent
- Building managed agents
- Environments
- Hooks
- Deep Research Agent
-
Tools
- Google Search
- Google Maps
- Code execution
- URL context
- Computer Use
- File Search
- Combine Tools and Function calling
-
Live API
-
Get started
- Live Translation
- Tool use
- Session management
- Ephemeral tokens
- Best practices
-
Optimization
- Batch API
- Webhooks
- Flex inference
- Priority inference
- Context caching
-
Guides
- Streaming
- Background execution
-
File input
- Media resolution
- Token counting
- Prompt engineering
-
Logs and datasets
-
Safety
-
Frameworks
-
Resources
- Deprecations
- Libraries
-
Migration
- Billing info
- API troubleshooting
- API errors
- Status
- Partner and library integrations
-
Google AI Studio
-
Google Cloud Platform
-
Policies
- Available regions
- Abuse monitoring
- Feedback information
The Interactions API is now generally available. We recommend using this API for access to all the latest features and models.
Send feedback
Models#
This guide introduces all the models available through the Gemini API.
Gemini 3#
Stable#
spark Gemini 3.6 Flash Our latest model that balances speed with intelligence to deliver strong performance in agentic and multimodal tasks. Stable spark Gemini 3.5 Flash Most intelligent model for sustained frontier performance on agentic and coding tasks. Stable bolt Gemini 3.5 Flash-Lite Our fastest, most cost-effective 3.5 model for high-throughput execution. Stable bolt Gemini 3.1 Flash-Lite Frontier-class performance rivaling larger models at a fraction of the cost. Stable 🍌🍌 Nano Banana 2 Powerful, high-efficiency image generation and editing, optimized for speed and high-volume use cases. Stable 🍌 Nano Banana 2 Lite Ultra-low latency and cost-effective image generation and editing, designed for high-volume interactive use cases. Stable 🍌 Nano Banana Pro State-of-the-art image generation and editing models for highly contextual native image creation. Stable
Preview#
auto_awesome Gemini 3.1 Pro Advanced intelligence, complex problem-solving skills, and powerful agentic and vibe coding capabilities. Preview spark Gemini 3 Flash Frontier-class performance rivaling larger models at a fraction of the cost. Preview translate Gemini 3.5 Live Translate Low-latency, real-time speech to speech translation model that supports 70+ languages. New Preview settings_voice Gemini 3.1 Flash Live High-quality, low-latency Live API model for real-time dialogue and voice-first AI applications. New Preview graphic_eq Gemini 3.1 Flash TTS Powerful, low-latency speech generation. New Preview movie_filter Gemini Omni Flash Fast, conversational video generation and editing. Turn text and images into video, and refine results through natural language. New Preview
All Gemini 3 models#
| Model | Endpoint |
|---|---|
| Gemini 3.6 Flash |
gemini-3.6-flash
gemini-3.5-flash
gemini-3.5-flash-lite
gemini-3.1-flash-lite
gemini-3.1-flash-image
gemini-3.1-flash-lite-image
gemini-3-pro-image
gemini-3.1-pro-preview
gemini-3-flash-preview
gemini-3.5-live-translate-preview
gemini-3.1-flash-live-preview
gemini-3.1-flash-tts-preview
gemini-omni-flash
Gemini 2.5 Flash#
| Model | Description | Endpoint |
|---|---|---|
| Gemini 2.5 Flash | Our best price-performance model for low-latency, high-volume tasks that require reasoning. |
gemini-2.5-flash
Nano Banana | State-of-the-art native image generation and editing designed for fast, creative workflows. |
gemini-2.5-flash-image
Gemini 2.5 Flash Live | Optimized for real-time conversational agents with sub-second native audio streaming. |
gemini-2.5-flash-native-audio-preview-12-2025
Gemini 2.5 Flash TTS | Controllable text-to-speech audio generation with fine control over style and pacing. |
gemini-2.5-flash-preview-tts
Gemini 2.5 Flash-Lite#
| Model | Description | Endpoint |
|---|---|---|
| Gemini 2.5 Flash-Lite | The fastest and most budget-friendly multimodal model in the 2.5 family. |
gemini-2.5-flash-lite
Gemini 2.5 Pro#
| Model | Description | Endpoint |
|---|---|---|
| Gemini 2.5 Pro | Our most advanced model for complex tasks, featuring deep reasoning and coding capabilities in the 2.5 family. |
gemini-2.5-pro
Gemini 2.5 Pro TTS | High-fidelity speech synthesis optimized for quality in structured workflows like podcasts and audiobooks. |
gemini-2.5-pro-preview-tts
Audio models#
This section contains all audio models, including ones that may already be listed in other sections
| Model | Description | Endpoint |
|---|---|---|
| Gemini 3.1 Flash Live | Our high-quality, low-latency audio-to-audio (A2A) model designed for real-time dialogue and voice-first AI applications. |
gemini-3.1-flash-live-preview
Gemini 3.1 Flash TTS | Powerful, low-latency speech generation, with natural outputs, steerable prompts, and new expressive audio tags for precise narration control. |
gemini-3.1-flash-tts-preview
Gemini 2.5 Flash Live | Our flagship Live API model for low-latency, bidirectional voice and video agents with native audio reasoning. |
gemini-2.5-flash-native-audio-preview-12-2025
Gemini 2.5 Flash TTS | Fast and controllable text-to-speech for low-latency, cost-efficient applications and real-time assistants. |
gemini-2.5-flash-preview-tts
Gemini 2.5 Pro TTS | High-fidelity speech synthesis optimized for quality in str