Gemini Developer API pricing
发布时间:2026-09-11 | 浏览:1
Español – América Latina
Português – Brasil
Start building free of charge with generous limits, then scale up with prepaid then pay-as-you-go pricing for your production ready applications.
For developers and small projects getting started with the Gemini API.
check_circle Limited access to certain models
check_circle Free input & output tokens
check_circle Google AI Studio access
check_circle Content used to improve our products *
For production applications that require higher volumes and advanced features.
check_circle Higher rate limits for production deployments
check_circle Access to Context caching
check_circle Batch API (50% cost reduction)
check_circle Access to Google's most advanced models
check_circle Content not used to improve our products *
For large-scale deployments with custom needs for security, support, and compliance, powered by Gemini Enterprise Agent Platform .
check_circle All features in Paid, plus optional access to:
check_circle Dedicated support channels
check_circle Advanced security & compliance
check_circle Provisioned throughput
check_circle Volume-based discounts (based on usage)
check_circle ML ops, model garden and more
Gemini 3.8 Flash
Try it in Google AI Studio
Our most intelligent Flash model, engineered for long-horizon software engineering, autonomous agents, and complex enterprise workflows.
Gemini 3.7 Flash
Try it in Google AI Studio
Our high-speed, efficient Flash model built for everyday coding, agentic tool use, and reliable multi-step execution.
Gemini 3.6 Flash
Try it in Google AI Studio
Our previous generation Flash model, balancing speed and multimodal capabilities across general agentic and everyday tasks.
* A customer-submitted request to Gemini may result in one or more queries to Google Search. You will be charged for each individual search query performed.
** Can be tested in Google AI Studio.
Gemini 3.5 Flash
Try it in Google AI Studio
Our earlier Flash model, built for speed and foundational performance across routine, high-throughput workloads.
* A customer-submitted request to Gemini may result in one or more queries to Google Search. You will be charged for each individual search query performed.
** Can be tested in Google AI Studio.
Gemini 3.5 Live Translate
Try it in Google AI Studio
Our low-latency, real-time speech to speech translation model that supports 70+ languages.
* Billing is based on total input and output audio token consumption, calculated at a rate of 25 tokens per second of audio, equating to an effective price of approximately $0.0368 per minute.
Gemini 3.5 Transcribe Live
Try it in Google AI Studio
Our low-latency, real-time speech-to-text model for bidirectional streaming audio transcription over WebSockets.
* Estimated pricing is based on 25 audio tokens per second for input and 175 text tokens per minute for output, for an effective blended rate of ~$0.009 per min for Live Transcribe.
Gemini 3.5 Transcribe
Try it in Google AI Studio
Our speech-to-text model with automatic language detection, speaker diarization, word-level timestamps, and custom vocabulary biasing.
* Estimated pricing is based on 25 audio tokens per second for input and 175 text tokens per minute for output, for an effective blended rate of ~$0.005 per min for Transcribe.
Gemini 3.5 Flash-Lite
Try it in Google AI Studio
A cost-efficient model, optimized for high-volume agentic tasks, translation, and simple data processing.
* A customer-submitted request to Gemini may result in one or more queries to Google Search. You will be charged for each individual search query performed.
** Can be tested in Google AI Studio.
** Can be tested in Google AI Studio.
Gemini 3.1 Flash-Lite
Try it in Google AI Studio
A cost-efficient model, optimized for high-volume agentic tasks, translation, and simple data processing.
* A customer-submitted request to Gemini may result in one or more queries to Google Search. You will be charged for each individual search query performed.
** Can be tested in Google AI Studio.
Gemini Omni Flash
Try it in Google AI Studio
Our next-generation video generation and editing model, now generally available to developers on the paid tier of the Gemini API.
* Billing is based on total output token consumption, calculated at a rate of 5,792 tokens per second of 720p video. Under Standard pricing, this equates to an effective price of approximately $0.10 per second.
Gemini Omni Flash Preview
Try it in Google AI Studio
Our next-generation video generation and editing model.
* Billing is based on total output token consumption, calculated at a rate of 5,792 tokens per second of 720p video. Under Standard pricing, this equates to an effective price of approximately $0.10 per second.
Gemini 3.1 Pro Preview
Try it in Google AI Studio
Our 3rd generation Pro model, built for multimodal understanding, agentic capabilities, and vibe-coding.
* A customer-submitted request to Gemini may result in one or more queries to Google Search. You will be charged for each individual search query performed.
** Can be tested in Google AI Studio.
Gemini 3.1 Flash Live Preview
Try it in Google AI Studio
Our low-latency, audio-to-audio model optimized for real-time dialogue with acoustic nuance detection, numeric precision, and multimodal awareness.
* A customer-submitted request to Gemini may result in one or more queries to Google Search. You will be charged for each individual search query performed.
Gemini 3.1 Flash Image (Nano Banana 2) 🍌
Try it in Google AI Studio
Designed for speed and efficiency, the Gemini 3.1 Flash Image generation model is effective for quick, interactive responses and high throughput.
* Image output is priced at $60 per 1,000,000 tokens. Output images at 0.5K (512px) consume 747 tokens and are equivalent to $0.045 per image. Output images at 1K (1024x1024px) consume 1120 tokens and are equivalent to $0.067 per image. Output images at 2K (2048x2048px) consume 1680 tokens and are equivalent to $0.101 per image. Output images at 4K (4096x4096px) consume 2520 tokens and are equivalent to $0.151 per image.
** A customer-submitted request to Gemini may result in one or more queries to Google Search. You will be charged for each individual search query performed. Retrieved context (text or images) provided by Grounding with Google Search is not charged as input tokens.
*** Can be tested in Google AI Studio.
Gemini 3.1 Flash Lite Image (Nano Banana 2 Lite) 🍌
Try it in Google AI Studio
Designed as the efficiency specialist of the image generation family, the Gemini 3.1 Flash Lite Image model is designed for ultra-low latency and cost-effective image generation and editing.
* Image output is priced at $30 per 1,000,000 tokens. Output images at 1K (1024x1024px) consume 1120 tokens and are equivalent to $0.0336 per image.
Gemini 3.1 Flash TTS Preview
Try it in Google AI Studio
Our 3.1 Flash Text-to-Speech audio model optimized for price-performant, low-latency, controllable speech generation.
* Audio tokens correspond to 25 tokens per second of audio.
Gemini 3 Flash Preview
Try it in Google AI Studio
Our legacy Flash model, providing baseline speed and intelligence.
* A customer-submitted request to Gemini may result in one or more queries to Google Search. You will be charged for each individual search query performed.
** Can be tested in Google AI Studio.
Gemini 3 Pro Image (Nano Banana Pro) 🍌
Try it in Google AI Studio
Our native image generation model, optimized for speed, flexibility, and contextual understanding. Text input and output is priced the same as Gemini 3.1 Pro .
* Image input is set at 560 tokens or $0.0011 per image.
** Image output is priced at $120 per 1,000,000 tokens. Output images from 1024x1024px (1K) and up to 2048x2048px (2K) consume 1120 tokens and are equivalent to $0.134 per image. Output images up to 4096x4096px (4K) consume 2000 tokens and are equivalent to $0.24 per image.
*** A customer-submitted request to Gemini may result in one or more queries to Google Search. You will be charged for each individual search query performed.
**** Can be tested in Google AI Studio.
Try it in Google AI Studio
A Pro model which excels at coding and complex reasoning tasks.
Gemini 2.5 Flash
Try it in Google AI Studio
Our first hybrid reasoning model which supports a 1M token context window and has thinking budgets.
Gemini 2.5 Flash-Lite
Try it in Google AI Studio
A small and cost effective model, built for at scale usage.
Gemini 2.5 Flash Native Audio (Live API)
Try it in Google AI Studio
Our Live API native audio models optimized for higher quality audio outputs with better pacing, voice naturalness, verbosity, and mood.
Gemini 2.5 Flash Image (Nano Banana) 🍌
Try it in Google AI Studio
A native image generation model, optimized for speed, flexibility, and contextual understanding. Text input and output is priced the same as 2.5 Flash .
[*] Image output is priced at $30 per 1,000,000 tokens. Output images up to 1024x1024px consume 1290 tokens and are equivalent to $0.039 per image.
Gemini 2.5 Flash Preview TTS
Try it in Google AI Studio
Our 2.5 Flash text-to-speech audio model optimized for price-performant, low-latency, controllable speech generation.
Gemini 2.5 Pro Preview TTS
Try it in Google AI Studio
Our 2.5 Pro text-to-speech audio model optimized for powerful, low-latency speech generation for more natural outputs and easier to steer prompts.
A fast video generation model.
Google's music generation model.
Google's family of legacy music generation models.
Gemini Embedding 2
Our first multimodal embedding model, mapping text, images, video, audio, and PDFs into a unified embedding space.
Gemini Embedding
Our first Gemini Embeddings model for text-only use cases.
Gemini Robotics ER 2 Preview
Try it in Google AI Studio
Gemini Robotics ER 2, short for Gemini Robotics Embodied Reasoning 2, is a vision-language model endpoint that enables robots to understand their environments precisely, supporting agentic orchestration of robots, video progress understanding, multi-robot collaboration, and advanced spatial reasoning.
Gemini Robotics ER 2 Streaming Preview
Try it in Google AI Studio
Gemini Robotics ER 2 Streaming is a vision-language model endpoint for robotics optimized for real-time text streaming using the Live API. It accepts text, image, video, and audio input and supports bidirectional streaming with function calling.
Gemini Robotics ER 1.6 Preview
Try it in Google AI Studio
Gemini Robotics ER, short for Gemini Robotics-Embodied Reasoning, is a thinking model that enhances robots' abilities to understand and interact with the physical world.
Gemini 2.5 Computer Use Preview
Our Computer Use model optimized for building browser control agents that automate tasks.
Our lightweight, state-of the art, open model built from the same technology that powers our Gemini models.
Pricing for tools
Tools are priced at their own rates, applied to the model using them. Check the Models page for which tools are available to each model.
Pricing for agents
Agent usage costs are calculated based on the underlying token consumption and usage of the tools.
Agentic video understanding: When using agentic video understanding, token usage is variable based on the content loaded by the model rather than full video length. This typically results in up to 88% fewer input tokens for long-form video, though token counts depend on query complexity and dynamic sampling depth (which may exceed 1 FPS for detailed visual segments). See Agentic video understanding .
Document token billing: Tokens for the DOCUMENT modality (for example, PDFs) are billed at the image token rate. In API responses, these tokens appear under the DOCUMENT modality within promptTokensDetails .
Google AI Studio usage is free of charge in all available regions . See Billing FAQs for details.
Prices may differ from the prices listed here and the prices offered on Gemini Enterprise Agent Platform. For Gemini Enterprise Agent Platform prices, see the Gemini Enterprise Agent Platform pricing page .
If you are using dynamic retrieval to optimize costs, only requests that contain at least one grounding support URL from the web in their response are charged for Grounding with Google Search. Costs for Gemini always apply. Rate limits are subject to change.
Except as otherwise noted, the content of this page is licensed under the Creative Commons Attribution 4.0 License , and code samples are licensed under the Apache 2.0 License . For details, see the Google Developers Site Policies . Java is a registered trademark of Oracle and/or its affiliates.
Last updated 2026-09-08 UTC.