Gemini Developer API pricing
发布时间:2026-08-25 | 浏览:1
Español – América Latina
Português – Brasil
Start building free of charge with generous limits, then scale up with prepaid then pay-as-you-go pricing for your production ready applications.
For developers and small projects getting started with the Gemini API.
check_circle Limited access to certain models
check_circle Free input & output tokens
check_circle Google AI Studio access
check_circle Content used to improve our products *
For production applications that require higher volumes and advanced features.
check_circle Higher rate limits for production deployments
check_circle Access to Context caching
check_circle Batch API (50% cost reduction)
check_circle Access to Google's most advanced models
check_circle Content not used to improve our products *
For large-scale deployments with custom needs for security, support, and compliance, powered by Gemini Enterprise Agent Platform .
check_circle All features in Paid, plus optional access to:
check_circle Dedicated support channels
check_circle Advanced security & compliance
check_circle Provisioned throughput
check_circle Volume-based discounts (based on usage)
check_circle ML ops, model garden and more
Gemini 3.7 Flash
Try it in Google AI Studio
Our most capable Flash model for agentic workflows and multimodal reasoning.
Gemini 3.6 Flash
Try it in Google AI Studio
Our most intelligent model built for speed, combining frontier intelligence with superior search and grounding.
* A customer-submitted request to Gemini may result in one or more queries to Google Search. You will be charged for each individual search query performed.
** Can be tested in Google AI Studio.
Gemini 3.5 Flash
Try it in Google AI Studio
Our most intelligent model built for speed, combining frontier intelligence with superior search and grounding.
* A customer-submitted request to Gemini may result in one or more queries to Google Search. You will be charged for each individual search query performed.
** Can be tested in Google AI Studio.
Gemini 3.5 Live Translate
Try it in Google AI Studio
Our low-latency, real-time speech to speech translation model that supports 70+ languages.
* Billing is based on total input and output audio token consumption, calculated at a rate of 25 tokens per second of audio, equating to an effective price of approximately $0.0368 per minute.
Gemini 3.5 Flash-Lite
Try it in Google AI Studio
Our most cost-efficient GA model, optimized for high-volume agentic tasks, translation, and simple data processing.
* A customer-submitted request to Gemini may result in one or more queries to Google Search. You will be charged for each individual search query performed.
** Can be tested in Google AI Studio.
** Can be tested in Google AI Studio.
Gemini 3.1 Flash-Lite
Try it in Google AI Studio
Our most cost-efficient model, optimized for high-volume agentic tasks, translation, and simple data processing.
* A customer-submitted request to Gemini may result in one or more queries to Google Search. You will be charged for each individual search query performed.
** Can be tested in Google AI Studio.
Gemini Omni Flash Preview
Try it in Google AI Studio
Our next-generation video generation and editing model, available to developers on the paid tier of the Gemini API.
* Billing is based on total output token consumption, calculated at a rate of 5,792 tokens per second of 720p video. Under Standard pricing, this equates to an effective price of approximately $0.10 per second.
Gemini 3.1 Pro Preview
Try it in Google AI Studio
The latest performance, intelligence, and usability improvements to the best model family in the world for multimodal understanding, agentic capabilities, and vibe-coding.
* A customer-submitted request to Gemini may result in one or more queries to Google Search. You will be charged for each individual search query performed.
** Can be tested in Google AI Studio.
Gemini 3.1 Flash Live Preview
Try it in Google AI Studio
Our low-latency, audio-to-audio model optimized for real-time dialogue with acoustic nuance detection, numeric precision, and multimodal awareness.
* A customer-submitted request to Gemini may result in one or more queries to Google Search. You will be charged for each individual search query performed.
Gemini 3.1 Flash Image (Nano Banana 2) 🍌
Try it in Google AI Studio
Designed for speed and efficiency, the Gemini 3.1 Flash Image generation model is effective for quick, interactive responses and high throughput.
* Image output is priced at $60 per 1,000,000 tokens. Output images at 0.5K (512px) consume 747 tokens and are equivalent to $0.045 per image. Output images at 1K (1024x1024px) consume 1120 tokens and are equivalent to $0.067 per image. Output images at 2K (2048x2048px) consume 1680 tokens and are equivalent to $0.101 per image. Output images at 4K (4096x4096px) consume 2520 tokens and are equivalent to $0.151 per image.
** A customer-submitted request to Gemini may result in one or more queries to Google Search. You will be charged for each individual search query performed. Retrieved context (text or images) provided by Grounding with Google Search is not charged as input tokens.
*** Can be tested in Google AI Studio.
Gemini 3.1 Flash Lite Image (Nano Banana 2 Lite) 🍌
Try it in Google AI Studio
Designed as the efficiency specialist of the image generation family, the Gemini 3.1 Flash Lite Image model is designed for ultra-low latency and cost-effective image generation and editing.
* Image output is priced at $30 per 1,000,000 tokens. Output images at 1K (1024x1024px) consume 1120 tokens and are equivalent to $0.0336 per image.
Gemini 3.1 Flash TTS Preview
Try it in Google AI Studio
Our 3.1 Flash Text-to-Speech audio model optimized for price-performant, low-latency, controllable speech generation.
Preview models may change before becoming stable and have more restrictive rate limits.
* Audio tokens correspond to 25 tokens per second of audio.
Gemini 3 Flash Preview
Try it in Google AI Studio
Our most intelligent model built for speed, combining frontier intelligence with superior search and grounding.
* A customer-submitted request to Gemini may result in one or more queries to Google Search. You will be charged for each individual search query performed.
** Can be tested in Google AI Studio.
Gemini 3 Pro Image (Nano Banana Pro) 🍌
Try it in Google AI Studio
Our native image generation model, optimized for speed, flexibility, and contextual understanding. Text input and output is priced the same as Gemini 3.1 Pro .
* Image input is set at 560 tokens or $0.0011 per image.
** Image output is priced at $120 per 1,000,000 tokens. Output images from 1024x1024px (1K) and up to 2048x2048px (2K) consume 1120 tokens and are equivalent to $0.134 per image. Output images up to 4096x4096px (4K) consume 2000 tokens and are equivalent to $0.24 per image.
*** A customer-submitted request to Gemini may result in one or more queries to Google Search. You will be charged for each individual search query performed.
**** Can be tested in Google AI Studio.
Try it in Google AI Studio
Our state-of-the-art multipurpose model, which excels at coding and complex reasoning tasks.
Gemini 2.5 Flash
Try it in Google AI Studio
Our first hybrid reasoning model which supports a 1M token context window and has thinking budgets.
Gemini 2.5 Flash-Lite
Try it in Google AI Studio
Our smallest and most cost effective model, built for at scale usage.
Gemini 2.5 Flash-Lite Preview
Try it in Google AI Studio
The latest model based on Gemini 2.5 Flash lite optimized for cost-efficiency, high throughput and high quality.
Gemini 2.5 Flash Native Audio (Live API)
Try it in Google AI Studio
Our Live API native audio models optimized for higher quality audio outputs with better pacing, voice naturalness, verbosity, and mood.
Preview models may change before becoming stable and have more restrictive rate limits.
Gemini 2.5 Flash Image (Nano Banana) 🍌
Try it in Google AI Studio
Our native image generation model, optimized for speed, flexibility, and contextual understanding. Text input and output is priced the same as 2.5 Flash .
Preview models may change before becoming stable and have more restrictive rate limits.
[*] Image output is priced at $30 per 1,000,000 tokens. Output images up to 1024x1024px consume 1290 tokens and are equivalent to $0.039 per image.
Gemini 2.5 Flash Preview TTS
Try it in Google AI Studio
Our 2.5 Flash text-to-speech audio model optimized for price-performant, low-latency, controllable speech generation.
Preview models may change before becoming stable and have more restrictive rate limits.
Gemini 2.5 Pro Preview TTS
Try it in Google AI Studio
Our 2.5 Pro text-to-speech audio model optimized for powerful, low-latency speech generation for more natural outputs and easier to steer prompts.
Preview models may change before becoming stable and have more restrictive rate limits.
Gemini 2.0 Flash
[*] Image output is priced at $30 per 1,000,000 tokens. Output images up to 1024x1024px consume 1290 tokens and are equivalent to $0.039 per image.
Gemini 2.0 Flash-Lite
Try it in Google AI Studio
Our latest image generation model, with significantly better text rendering and better overall image quality.
Preview models may change before becoming stable and have more restrictive rate limits.
Our latest video generation model, available to developers on the paid tier of the Gemini API.
Preview models may change before becoming stable and have more restrictive rate limits.
Our stable video generation model, available to developers on the paid tier of the Gemini API.
Our state-of-the-art video generation model, available to developers on the paid tier of the Gemini API.
Google's family of music generation models. Preview models may change before becoming stable and have more restrictive rate limits.
Gemini Embedding 2
Our first multimodal embedding model, mapping text, images, video, audio, and PDFs into a unified embedding space.
Gemini Embedding
Our Gemini Embeddings model for text-only use cases, available to developers on the free and paid tiers of the Gemini API.
Gemini Robotics ER 2 Preview
Try it in Google AI Studio
Gemini Robotics ER 2, short for Gemini Robotics Embodied Reasoning 2, is a vision-language model endpoint that enables robots to understand their environments precisely, supporting agentic orchestration of robots, video progress understanding, multi-robot collaboration, and advanced spatial reasoning.
Gemini Robotics ER 2 Streaming Preview
Try it in Google AI Studio
Gemini Robotics ER 2 Streaming is a vision-language model endpoint for robotics optimized for real-time text streaming using the Live API. It accepts text, image, video, and audio input and supports bidirectional streaming with function calling.
Gemini Robotics ER 1.6 Preview
Try it in Google AI Studio
Gemini Robotics ER, short for Gemini Robotics-Embodied Reasoning, is a thinking model that enhances robots' abilities to understand and interact with the physical world.
Gemini 2.5 Computer Use Preview
Our Computer Use model optimized for building browser control agents that automate tasks.
Our lightweight, state-of the art, open model built from the same technology that powers our Gemini models.
Pricing for tools
Tools are priced at their own rates, applied to the model using them. Check the Models page for which tools are available to each model.
Pricing for agents
Agent usage costs are calculated based on the underlying token consumption and usage of the tools.
Document token billing: Tokens for the DOCUMENT modality (for example, PDFs) are billed at the image token rate. In API responses, these tokens appear under the DOCUMENT modality within promptTokensDetails .
Google AI Studio usage is free of charge in all available regions . See Billing FAQs for details.
Prices may differ from the prices listed here and the prices offered on Gemini Enterprise Agent Platform. For Gemini Enterprise Agent Platform prices, see the Gemini Enterprise Agent Platform pricing page .
If you are using dynamic retrieval to optimize costs, only requests that contain at least one grounding support URL from the web in their response are charged for Grounding with Google Search. Costs for Gemini always apply. Rate limits are subject to change.
Except as otherwise noted, the content of this page is licensed under the Creative Commons Attribution 4.0 License , and code samples are licensed under the Apache 2.0 License . For details, see the Google Developers Site Policies . Java is a registered trademark of Oracle and/or its affiliates.
Last updated 2026-08-13 UTC.