Google Gemini
Google · source page ↗ · last checked Aug 24, 2026, 12:00 PM
Current pricing
● live captured Aug 24, 2026, 12:00 PM| Model / item | Input | Output |
|---|---|---|
| Gemini 3.1 Pro Preview | $2.00 / 1M tokens (prompts ≤200k), $4.00 / 1M tokens (>200k) | $12.00 / 1M tokens (≤200k), $18.00 / 1M tokens (>200k) · Batch 50% off ($1.00/$6.00 ≤200k). Context caching $0.20/1M (≤200k), $0.40/1M (>200k) + $4.50/1M/hr storage. Free tier available. |
| Gemini 3.5 Flash | $1.50 / 1M tokens | $9.00 / 1M tokens · Batch/Flex 50% off ($0.75/$4.50). Priority $2.70/$16.20. Context caching $0.15/1M + $1.00/1M/hr storage. Free tier available. |
| Gemini 3 Flash Preview | $0.50 / 1M tokens (text/image/video), $1.00 / 1M tokens (audio) | $3.00 / 1M tokens · Batch/Flex 50% off. Priority $0.90 (text/image/video)/$1.80 (audio) input, $5.40 output. Free tier available. |
| Gemini 3.1 Flash-Lite | $0.25 / 1M tokens (text/image/video), $0.50 / 1M tokens (audio) | $1.50 / 1M tokens · Batch/Flex 50% off. Priority $0.45 (text/image/video)/$0.90 (audio) input, $2.70 output. Free tier available. |
| Gemini 2.5 Pro | $1.25 / 1M tokens (prompts ≤200k), $2.50 / 1M tokens (prompts >200k) | $10.00 / 1M tokens (prompts ≤200k), $15.00 / 1M tokens (prompts >200k) · Batch/Flex 50% off ($0.625/$5.00 ≤200k). Priority $2.25/$18.00 (≤200k). Free tier available. |
| Gemini 2.5 Flash | $0.30 / 1M tokens (text/image/video), $1.00 / 1M tokens (audio) | $2.50 / 1M tokens · Batch/Flex 50% off ($0.15/$1.25). Priority $0.54 (text/image/video)/$1.80 (audio) input, $4.50 output. Context caching $0.03/1M + $1.00/1M/hr storage. Free tier available. |
| Gemini 2.5 Flash-Lite | $0.10 / 1M tokens (text/image/video), $0.30 / 1M tokens (audio) | $0.40 / 1M tokens · Batch/Flex 50% off ($0.05/$0.20). Priority $0.18 (text/image/video)/$0.54 (audio) input, $0.72 output. Free tier available. |
| Gemini 2.0 Flash (deprecated) | $0.10 / 1M tokens (text/image/video), $0.70 / 1M tokens (audio) | $0.40 / 1M tokens · Deprecated — shut down June 1, 2026. Batch 50% off. Context caching $0.025/1M + $1.00/1M/hr storage. Free tier available. |
| Gemini 2.0 Flash-Lite (deprecated) | $0.075 / 1M tokens | $0.30 / 1M tokens · Deprecated — shut down June 1, 2026. Batch 50% off ($0.0375/$0.15). Free tier available. |
| Gemini Embedding 2 | $0.20 / 1M tokens (text); image $0.45/1M; audio $6.50/1M; video $12.00/1M | N/A (embeddings model) · Batch 50% off (text $0.10/1M). Free tier available. |
| Gemini Embedding (001) | $0.15 / 1M tokens (text) | N/A (embeddings model) · Batch 50% off ($0.075/1M). Free tier available (text only). |
| Imagen 4 | Per-image pricing | $0.02/image (Fast), $0.04/image (Standard), $0.06/image (Ultra) · Paid tier only. |
| Veo 3.1 | Per-second of video | Standard $0.40/sec (720p/1080p), $0.60/sec (4K); Fast $0.10–$0.30/sec; Lite $0.05–$0.08/sec · Charged only on successful generation. Paid tier only. |
| Google Search Grounding (tool) | Per grounded query | Gemini 3: 5,000 free prompts/mo, then $14 / 1,000 queries. Gemini 2.5: 1,500 free RPD, then $35 / 1,000 grounded prompts · Retrieved context tokens excluded from billing. |
| Gemini 2.0 Flash | $0.75 / 1M tokens through December 31, 2026; $1.50 / 1M tokens starting January 1, 2027 | $3.75 / 1M tokens through December 31, 2026; $7.50 / 1M tokens starting January 1, 2027 |
| Gemini 2.0 Flash Lite | $0.375 / 1M tokens through December 31, 2026; $0.75 / 1M tokens starting January 1, 2027 | $1.875 / 1M tokens through December 31, 2026; $3.75 / 1M tokens starting January 1, 2027 |
| Gemini 2.0 Pro | $1.35 / 1M tokens through December 31, 2026; $2.70 / 1M tokens starting January 1, 2027 | $6.75 / 1M tokens through December 31, 2026; $13.50 / 1M tokens starting January 1, 2027 |
| Gemini 1.5 Pro | $0.75 / 1M tokens through December 31, 2026; $1.50 / 1M tokens starting January 1, 2027 | $3.75 / 1M tokens through December 31, 2026; $7.50 / 1M tokens starting January 1, 2027 |
| Gemini 1.5 Flash | $0.375 / 1M tokens through December 31, 2026; $0.75 / 1M tokens starting January 1, 2027 | $1.875 / 1M tokens through December 31, 2026; $3.75 / 1M tokens starting January 1, 2027 |
| Gemini 3 Pro | $0.50 / 1M tokens | $3.00 / 1M tokens |
| Gemini 3 Flash | $0.15 / 1M tokens | $1.25 / 1M tokens |
| Gemini 3.1 Pro | $2.00 / 1M tokens (prompts ≤200k), $4.00 / 1M tokens (prompts >200k) | $12.00 / 1M tokens (prompts ≤200k), $18.00 / 1M tokens (prompts >200k) |
| Gemini 3.1 Sonnet | $0.75 / 1M tokens (text), $3.00 / 1M tokens (audio) | $4.50 / 1M tokens (text), $12.00 / 1M tokens (audio) |
Usage/token-based API pricing (no subscription tiers). Prices in USD per 1M tokens unless noted. All models offer Standard, and most offer Batch (~50% off), Flex, and Priority service tiers. Many models have context length tiers (prompts ≤200k vs >200k tokens), separate audio rates, and context caching. Free tier (lower rate limits; data may be used to improve products) vs paid tier (data not used without consent). Enterprise/Vertex AI prices may differ. Preview models may change before stabilization. Page last updated June 9, 2026; prices subject to change. Source returned via WebFetch; additional models on page include Gemini 3 Pro Image, Gemini 2.5 Flash Image, TTS/Live audio models, Lyria 3 music, and Computer Use / Robotics specialized models.
Rate & usage limits ● livestructure only
This provider doesn't publish numeric request/token limits on its docs — only the usage-tier structure below. Your actual limits appear in your account dashboard.
| Tier | Notes |
|---|---|
| Free | Active project or free trial; no spend-based rate limits |
| Tier 1 | Set up billing account; $10 spend-based limit per 10 minutes; $250 billing tier cap |
| Tier 2 | Paid $100 + 3 days from first payment; $200 spend-based limit per 10 minutes; $2,000 billing tier cap |
| Tier 3 | Paid $1,000 + 30 days from first payment; $200 spend-based limit per 10 minutes; $20,000–$100,000+ billing tier cap |
Numeric RPM, TPM, and RPD limits are not published on this page; actual limits are only viewable per-project in AI Studio dashboard. Rate limits vary by model and are automatically updated based on usage tier. Priority inference gets 0.3x standard rates. Batch API has separate concurrent request (100), file size (2GB), and enqueued token limits per model and tier.
Deprecations
Gemini 3 models
Gemini 2.5 Pro models
Gemini 2.5 Flash models
Gemini 2.0 models
Live API models
Audio models
Embedding models
Imagen models
Veo models
Lyria models
Robotics models
Per the page: a "deprecation" is the announcement that a model is no longer supported and will be "shut down" soon; once "shutdown", the endpoint is completely turned off and no longer available. Shutdown dates are the earliest possible retirement dates; exact dates are communicated separately. Status here reflects the page's version label (GA / Preview / Experimental); items with a shutdown/retirement date are marked Deprecated, and those with "No shutdown announced" are Active. Dates printed exactly as on the page; release dates given as "Released On". Some preview/experimental rows have no release date listed on the page.
Curated snapshot · as of 2026-06-15. Detected page changes are in the change history below.
Changelog
- June 1, 2026 Gemini 2.0 models shut downFour Gemini 2.0 Flash models discontinued; users directed to Gemini 3.5 alternatives.
- May 28, 2026 Native visual models GA & video-to-image supportTwo native image generation models released as GA; video files now usable as context.
- May 25, 2026 Flash-Lite preview model shut downgemini-3.1-flash-lite-preview discontinued; GA version available as replacement.
- May 19, 2026 Gemini 3.5 Flash GA & managed agents previewGemini 3.5 Flash released for sustained frontier performance; autonomous sandboxed agents launched in preview.
- May 7, 2026 Flash-Lite GA & deprecation noticeCost-efficient Gemini 3.1 Flash-Lite released as stable; preview version deprecating May 25.
- May 6, 2026 Interactions API breaking change noticeRequest/response schema and output format configuration changing; migration guide provided.
- May 5, 2026 File Search multimodal supportFile Search now supports image embedding and search via the gemini-embedding-2 model.
- May 4, 2026 Webhooks support launchedEvent-driven webhooks replace polling for the Batch API and long-running operations.
- April 30, 2026 Robotics model shut downgemini-robotics-er-1.5-preview discontinued; newer 1.6 version available.
- April 22, 2026 Embedding model GAgemini-embedding-2 released as generally available.
- April 21, 2026 Deep Research agent updatedTwo new versions add collaborative planning, visualization, MCP, and File Search support.
- April 15, 2026 Text-to-speech model launchedGemini 3.1 Flash TTS Preview released as a cost-efficient speech generation model.
- April 14, 2026 Robotics model update & deprecationVersion 1.6 adds instrument reading and spatial reasoning; 1.5 retiring April 30.
- April 2, 2026 Gemma 4 models releasedTwo new Gemma 4 models available via AI Studio and the Gemini API.
- April 1, 2026 Flex & Priority inference tiers introducedNew inference tiers offering more options for optimizing cost or latency.
- March 31, 2026 Veo 3.1 Lite & model shutdownCost-efficient video generation model released; older Flash-Lite model discontinued.
- March 26, 2026 Audio-to-audio model releasedLatest A2A model for real-time dialogue and voice-first AI applications launched.
- March 25, 2026 Lyria 3 music generation launchTwo music models released supporting text and image inputs; 30-second and full-length options.
- March 23, 2026 Billing plans rolloutPrepay and Postpay billing options introduced in AI Studio.
- March 18, 2026 Tools & function combination featureBuilt-in tools now work alongside custom function calling in a single API request.
- March 16, 2026 Usage tiers & spend caps revampedRevamped billing structure with updated usage tiers and account spend limits.
- March 12, 2026 Project-level spend caps addedProject-specific spending limits now available in billing configuration.
- March 10, 2026 Multimodal embedding model & deprecationFirst multimodal embedding model supports five modalities; older embedding model retiring.
- March 9, 2026 Gemini 3 Pro redirectOriginal Gemini 3 Pro Preview shut down; model name now aliases to the 3.1 version.
- March 3, 2026 Flash-Lite preview launchedFirst Flash-Lite model in the Gemini 3 series released in preview.
- February 26, 2026 Nano Banana 2 & deprecation noticeHigh-efficiency image generation model released; older Pro version retiring March 9.
- February 19, 2026 Gemini 3.1 Pro & custom tools endpointLatest iteration released with new media resolution behavior; specialized tools endpoint added.
- February 18, 2026 Gemini 2.0 models deprecation announcedFour Flash models scheduled for June 1 shutdown; replacement models recommended.
- February 17, 2026 Model shutdownsThree models discontinued; users directed to newer alternatives.
- January 29, 2026 Computer Use tool supportComputer Use tool now supported on two Gemini 3 preview models.
- January 21, 2026 Latest alias updatesModel aliases updated to point to Gemini 3 preview versions.
- January 15, 2026 Model shutdowns & deprecationsThree models retiring February 17; future deprecations announced.
- January 14, 2026 Embedding model shutdowntext-embedding-004 discontinued.
- January 13, 2026 Video output resolution enhancement4K output added for Veo; portrait video support expanded across resolutions.
- January 12, 2026 Model lifecycle feature launchedModels now specify lifecycle stage and deprecation timeline for clarity.
- January 8, 2026 File input expansionCloud Storage buckets and pre-signed URLs supported; file size limit raised to 100MB.
- December 19, 2025 Interactions API field renametotal_reasoning_tokens renamed to total_thought_tokens in a breaking change.
- December 17, 2025 Gemini 3 Flash Preview releasedFast frontier-class model launched with upgraded reasoning and new multimodal features.
- December 12, 2025 Native audio model releasedNew audio model for the Live API with improved complex workflow handling.
- December 11, 2025 Interactions API beta & Deep Research agentUnified API interface launched; autonomous research agent in preview.
Google Gemini API release notes. Page last updated 2026-06-01 UTC. Entries are summarized from the official changelog; dates and model/version names are reproduced as printed on the page.
Curated snapshot · as of 2026-06-15. Detected page changes are in the change history below.
In the news · Google
- Google Leaks Gemini 3.5 Pro Launching Today: Here Is What Changes Everything
- Best Google Gemini features to use for free in Nigeria right now
- I switched to Firefox, but Google's Gemini pulled me back to Chrome
- Hackers Use Fake Google Gemini Installer to Deploy Vidar Stealer and Steal Browser Credentials
- Free Google Gemini for Students: How to Claim Your Subscription - hi-Tech.ua
- Hackers Use Fake Google Gemini App to Steal Windows Users’ Browser Credentials
- College AI Communities Are Taking Off: Everytime Launches Dedicated Google Gemini Lounge
- Hackers Use Fake Google Gemini Installer to Deploy Vidar Stealer and Steal Browser Credentials
Importance-filtered press coverage (Google News) mentioning Google. Headlines link to the original; verify before acting.
Change history
- Update Aug 14, 2026, 12:00 AM
Google Gemini deprecations changed
Added gemini-3.7-flash model with August 2026 release date and no announced shutdown date.
- gemini-3.7-flash
- August 2026
- No shutdown date announced
View raw diff +1 −0
+ | gemini-3.7-flash | August 2026 | No shutdown date announced | | - Pricing Aug 13, 2026, 6:00 PM
Google Gemini API pricing changed
Gemini API pricing increases announced for January 1, 2027, with rates doubling for multiple models including Flash, Pro, and Sonnet variants.
Gemini models now show dated pricing tiers: current rates through Dec 31, 2026, then doubled rates starting Jan 1, 2027View raw diff +64 −4
- $7.50 - $3.75 - $3.75 - $13.50 + $0.75 through December 31, 2026. + $1.50 starting January 1, 2027. + $3.75 through December 31, 2026. + $7.50 starting January 1, 2027. + $0.075 through December 31, 2026. + $0.15 starting January 1, 2027. + $0.50 / 1,000,000 tokens per hour (storage price) through December 31, 2026. + $1.00 / 1,000,000 tokens per hour (storage price) starting January 1, 2027. + $0.375 through December 31, 2026. + $0.75 starting January 1, 2027. + $1.875 through December 31, 2026. + $3.75 starting January 1, 2027. + $0.0375 through December 31, 2026. + $0.075 starting January 1, 2027. + $0.50 / 1,000,000 tokens per hour (storage price) through December 31, 2026. + $1.00 / 1,000,000 tokens per hour (storage price) starting January 1, 2027. + $0.375 through December 31, 2026. + $0.75 starting January 1, 2027. + $1.875 through December 31, 2026. + $3.75 starting January 1, 2027. + $0.0375 through December 31, 2026. + $0.075 starting January 1, 2027. + $0.50 / 1,000,000 tokens per hour (storage price) through December 31, 2026. + $1.00 / 1,000,000 tokens per hour (storage price) starting January 1, 2027. + $1.35 through December 31, 2026. + $2.70 starting January 1, 2027. + $6.75 through December 31, 2026. + $13.50 starting January 1, 2027. + $0.135 through December 31, 2026. + $0.27 starting January 1, 2027. + $0.50 / 1,000,000 tokens per hour (storage price) through December 31, 2026. + $1.00 / 1,000,000 tokens per hour (storage price) starting January 1, 2027. + $0.75 through December 31, 2026. + $1.50 starting January 1, 2027. + $3.75 through December 31, 2026. + $7.50 starting January 1, 2027. + $0.075 through December 31, 2026. + $0.15 starting January 1, 2027. + $0.50 / 1,000,000 tokens per hour (storage price) through December 31, 2026. + $1.00 / 1,000,000 tokens per hour (storage price) starting January 1, 2027. + $0.375 through December 31, 2026. + $0.75 starting January 1, 2027. + $1.875 through December 31, 2026. + $3.75 starting January 1, 2027. + $0.0375 through December 31, 2026. + $0.075 starting January 1, 2027. + $0.50 / 1,000,000 tokens per hour (storage price) through December 31, 2026. + $1.00 / 1,000,000 tokens per hour (storage price) starting January 1, 2027. + $0.375 through December 31, 2026. + $0.75 starting January 1, 2027. + $1.875 through December 31, 2026. + $3.75 starting January 1, 2027. + $0.0375 through December 31, 2026. + $0.075 starting January 1, 2027. + $0.50 / 1,000,000 tokens per hour (storage price) through December 31, 2026. + $1.00 / 1,000,000 tokens per hour (storage price) starting January 1, 2027.
- Update Aug 11, 2026, 6:00 AM
Google Gemini API changelog changed
Multiple model releases to GA and preview including Gemini 3.6 Flash, Gemini 3.5 Flash, Gemini 3.1 Flash-Lite, image/video generation models, and new Managed Agents with Antigravity Agent, plus several deprecation announcements with shutdown dates.
- Gemini 3.5 Flash and Gemini 3.6 Flash released to GA, gemini-3.5-flash is now the model behind gemini-flash-latest
- Managed Agents in Gemini API launched in public preview with Antigravity Agent (antigravity-preview-05-2026)
- Multiple deprecation announcements: gemini-robotics-er-1.6-preview shuts down August 31 2026, image generation models August 17 2026, video generation models June 30 2026
View raw diff +66 −192
- Gemini Robotics ER 2 in public preview: Released two new embodied - reasoning model endpoints for robotics: - Both model endpoints accept text, image, video, and audio inputs and support - function calling with blocking behavior for physical robot actions. - To get started, see the - Gemini Robotics ER overview. For - real-time streaming use cases, see - Robotics with streaming. - Deprecation announcement: The gemini-robotics-er-1.6-preview model - will be shut down on August 31, 2026. - Gemini 3.6 Flash and Gemini 3.5 Flash-Lite generally available (GA): - Released stable, production-ready versions of our latest 3.x Flash models: - To learn more, see the Latest Gemini model - guide. - Deprecated parameters: The sampling parameters temperature, top_p - and top_k are now deprecated. See the - Latest Gemini Model - for details. - Gemini Omni Flash in public preview: Released gemini-omni-flash-preview, - a high-performance multimodal model designed for high-speed video generation - and conversational video editing. Using the Interactions API, - you can generate 3--10 second videos at 720p from text descriptions or animate still images, - and then conversationally edit and refine the outputs. To get started, see the - Gemini Omni Flash guide and the - Gemini Omni Flash model card. - Released gemini-3.1-flash-lite-image (Nano Banana 2 Lite) to general - availability (GA), our built-in multimodal model optimized for ultra-low - latency and cost-effective image generation and editing. See the [Gemini 3.1 - Flash Lite Image](https://ai.google.dev/gemini-api/docs/models/gemini-3.1-flash-lite-image) model - card and the Image generation guide. - Deprecation announcement: The following image generation models are - being deprecated and will be shut down on August 17, 2026: - To migrate your code to newer stable or preview endpoints, refer to the - Gemini deprecations page. - Deprecation announcement: The following video generation models are - being deprecated and will be shut down on June 30, 2026: - Update your integration to either use the Veo 3.1 preview model IDs - (veo-3.1-generate-preview, veo-3.1-fast-generate-preview) or the - 3.1 GA models available through the - Gemini Enterprise Agent Platform - to avoid service interruptions. - Use gemini-3.5-flash or - gemini-3.1-flash-lite - instead. - Released gemini-3.1-flash-image (Nano Banana 2) and gemini-3-pro-image - (Nano Banana Pro), the generally available (GA) versions of our native - visual models, Gemini 3.1 Flash Image - and Gemini 3 Pro Image. - Video-to-image generation support: You can now pass a video file (via - direct upload or as a public YouTube URL) as multimodal context alongside a - text prompt to generate high-quality thumbnails, cinematic movie posters, or - summary infographics. This feature is supported exclusively on the - gemini-3.1-flash-image model. To learn more, see the - Video-to-image generation - guide. - Deprecation announcement: The gemini-3.1-flash-image-preview and - gemini-3-pro-image-preview models are deprecated - and will be shut down on June 25, 2026. - Released gemini-3.5-flash, the generally available (GA) version of - Gemini 3.5 Flash,
- Update Aug 3, 2026, 6:00 PM
Google Gemini deprecations changed
Gemini 2.5 model variants no longer have announced shutdown dates, removing the October 16, 2026 deadline.
- gemini-2.5-pro shutdown date changed from October 16, 2026 to No shutdown date announced
- gemini-2.5-flash shutdown date changed from October 16, 2026 to No shutdown date announced
- gemini-2.5-flash-lite shutdown date changed from October 16, 2026 to No shutdown date announced
View raw diff +3 −3
- | gemini-2.5-pro | June 17, 2025 | October 16, 2026 | gemini-3.1-pro-preview | - | gemini-2.5-flash | June 17, 2025 | October 16, 2026 | gemini-3.6-flash | - | gemini-2.5-flash-lite | July 22, 2025 | October 16, 2026 | gemini-3.1-flash-lite | + | gemini-2.5-pro | June 17, 2025 | No shutdown date announced | | + | gemini-2.5-flash | June 17, 2025 | No shutdown date announced | | + | gemini-2.5-flash-lite | July 22, 2025 | No shutdown date announced | |
- Update Jul 30, 2026, 6:00 PM
Google Gemini API changelog changed
Released Gemini Robotics ER 2 with two new model endpoints (gemini-robotics-er-2-preview and gemini-robotics-er-2-streaming-preview) and deprecated gemini-robotics-er-1.6-preview effective August 31, 2026.
- gemini-robotics-er-2-preview: Advanced spatial reasoning with multi-robot coordination
- gemini-robotics-er-2-streaming-preview: Optimized for real-time text streaming with Live API
- gemini-robotics-er-1.6-preview will be shut down on August 31, 2026
View raw diff +15 −2
- Gemini Robotics-ER page and the - Released Gemini Robotics-ER 1.5 model in preview. See the + July 30, 2026 + Gemini Robotics ER 2 in public preview: Released two new embodied + reasoning model endpoints for robotics: + gemini-robotics-er-2-preview: Advanced spatial reasoning, agentic code execution, multi-step tool orchestration, video moment finding, progress classification, and multi-robot coordination. + gemini-robotics-er-2-streaming-preview: Optimized for real-time text streaming using the Live API, enabling low-latency robot agents with bidirectional audio and video input. + Both model endpoints accept text, image, video, and audio inputs and support + function calling with blocking behavior for physical robot actions. + To get started, see the + Gemini Robotics ER overview. For + real-time streaming use cases, see + Robotics with streaming. + Deprecation announcement: The gemini-robotics-er-1.6-preview model + will be shut down on August 31, 2026. + Gemini Robotics ER page and the + Released Gemini Robotics ER 1.5 model in preview. See the
- Update Jul 30, 2026, 6:00 PM
Google Gemini deprecations changed
Extended deprecation timeline for gemini-embedding-001 to May 14, 2028 and added shutdown date of August 31, 2026 for gemini-robotics-er-1.6-preview with upgrade path to gemini-robotics-er-2-preview.
- gemini-embedding-001 deprecation date extended from July 14, 2026 to May 14, 2028
- gemini-robotics-er-1.6-preview shutdown date set to August 31, 2026 (previously had no announced date)
- gemini-robotics-er-1.6-preview migration path added to gemini-robotics-er-2-preview
View raw diff +2 −2
- | gemini-embedding-001 | July 14, 2025 | July 14, 2026 | gemini-embedding-2 | - | gemini-robotics-er-1.6-preview | April 14, 2026 | No shut down date announced | | + | gemini-embedding-001 | July 14, 2025 | May 14, 2028 | gemini-embedding-2 | + | gemini-robotics-er-1.6-preview | April 14, 2026 | August 31, 2026 | gemini-robotics-er-2-preview |
- Pricing Jul 30, 2026, 6:00 PM
Google Gemini API pricing changed
Search pricing terms standardized across all Gemini models to unified monthly free request limits and cost structure
Search pricing terms unified to "5,000 free search requests per month (shared across all Gemini 3.x models), then $14 per 1,000 requests" across all modelsView raw diff +39 −3
- 5,000 prompts per month (free, shared across Gemini 3), then $14 / 1,000 search queries for text and image-based grounding. - 5,000 prompts per month (free, limit shared with Flash), then $14 / 1,000 search queries for text and image-based grounding. - 5,000 prompts per month (free), then $14 / 1,000 search queries + 5,000 free search requests per month (shared across all Gemini 3.x models), then $14 per 1,000 requests. + 5,000 free search requests per month (shared across all Gemini 3.x models), then $14 per 1,000 requests. + 5,000 free search requests per month (shared across all Gemini 3.x models), then $14 per 1,000 requests. + 5,000 free search requests per month (shared across all Gemini 3.x models), then $14 per 1,000 requests. + 5,000 free search requests per month (shared across all Gemini 3.x models), then $14 per 1,000 requests. + 5,000 free search requests per month (shared across all Gemini 3.x models), then $14 per 1,000 requests. + 5,000 free search requests per month (shared across all Gemini 3.x models), then $14 per 1,000 requests. + 5,000 free search requests per month (shared across all Gemini 3.x models), then $14 per 1,000 requests. + 5,000 free search requests per month (shared across all Gemini 3.x models), then $14 per 1,000 requests. + 5,000 free search requests per month (shared across all Gemini 3.x models), then $14 per 1,000 requests. + 5,000 free search requests per month (shared across all Gemini 3.x models), then $14 per 1,000 requests. + 5,000 free search requests per month (shared across all Gemini 3.x models), then $14 per 1,000 requests. + 5,000 free search requests per month (shared across all Gemini 3.x models), then $14 per 1,000 requests. + 5,000 free search requests per month (shared across all Gemini 3.x models), then $14 per 1,000 requests. + 5,000 free search requests per month (shared across all Gemini 3.x models), then $14 per 1,000 requests. + 5,000 free search requests per month (shared across all Gemini 3.x models), then $14 per 1,000 requests. + 5,000 free search requests per month (shared across all Gemini 3.x models), then $14 per 1,000 requests. + 5,000 free search requests per month (shared across all Gemini 3.x models), then $14 per 1,000 requests. + 5,000 free search requests per month (shared across all Gemini 3.x models), then $14 per 1,000 requests. + 5,000 free search requests per month (shared across all Gemini 3.x models), then $14 per 1,000 requests. + 5,000 free search requests per month (shared across all Gemini 3.x models), then $14 per 1,000 requests. + 5,000 free search requests per month (shared across all Gemini 3.x models), then $14 per 1,000 requests for text and image-based grounding. + 5,000 free search requests per month (shared across all Gemini 3.x models), then $14 per 1,000 requests. + 5,000 free search requests per month (shared across all Gemini 3.x models), then $14 per 1,000 requests. + 5,000 free search requests per month (shared across all Gemini 3.x models), then $14 per 1,000 requests. + 5,000 free search requests per month (shared across all Gemini 3.x models), then $14 per 1,000 requests. + 5,000 free search requests per month (shared across all Gemini 3.x models), then $14 per 1,000 requests. + 5,000 free search requests per month (shared across all Gemini 3.x models), then $14 per 1,000 requests. + $2.00 (text / image / video / audio) + $10.00 + 5,000 free search requests per month (shared across all Gemini 3.x models), then $14 per 1,000 requests. + $1.00 (text / image / video / audio) + 5,000 free search requests per month (shared across all Gemini 3.x models), then $14 per 1,000 requests. + $2.00 (text / image / video / audio) + $10.00 + 5,000 free search requests per month (shared across all Gemini 3.x models), then $14 per 1,000 requests. + 5,000 free search requests per month (shared across all Gemini 3.x models), then $14 per 1,000 requests. + 5,000 free search requests per month (shared across all Gemini 3.x models), then $14 per 1,000 requests. + 5,000 free search requests per month (shared across all Gemini models), then $14 per 1,000 requests.
- Update Jul 21, 2026, 6:00 PM
Google Gemini API changelog changed
Gemini 3.6 Flash and 3.5 Flash-Lite models released as GA; temperature, top_p, and top_k sampling parameters deprecated.
- Gemini 3.6 Flash (gemini-3.6-flash) available as stable production model with improved token efficiency and lower price
- Gemini 3.5 Flash-Lite (gemini-3.5-flash-lite) released as low-latency, cost-effective subagent option
- Sampling parameters temperature, top_p, and top_k are now deprecated
View raw diff +10 −0
+ July 21, 2026 + Gemini 3.6 Flash and Gemini 3.5 Flash-Lite generally available (GA): + Released stable, production-ready versions of our latest 3.x Flash models: + Gemini 3.6 Flash (gemini-3.6-flash): Features improved token efficiency and code/agentic planning capabilities at a lower price point than 3.5 Flash, resolving developer feedback around output verbosity. + Gemini 3.5 Flash-Lite (gemini-3.5-flash-lite): Offers a low-latency, highly cost-effective subagent option designed for high-volume automation. + To learn more, see the Latest Gemini model + Deprecated parameters: The sampling parameters temperature, top_p + and top_k are now deprecated. See the + Latest Gemini Model + for details.
- Update Jul 21, 2026, 6:00 PM
Google Gemini deprecations changed
Model deprecation guidance updated with two new models (gemini-3.6-flash and gemini-3.5-flash-lite) and migration targets changed from gemini-3.5-flash to gemini-3.6-flash for most older models.
- New model gemini-3.6-flash released July 21, 2026 with no shutdown date announced
- New model gemini-3.5-flash-lite released July 21, 2026 with no shutdown date announced
- Migration target for gemini-2.5-flash, gemini-2.0-flash, and related versions changed from gemini-3.5-flash to gemini-3.6-flash
View raw diff +9 −7
- | gemini-3.1-flash-lite | May 7, 2026 | May 7, 2027 | | - | gemini-3-flash-preview | December 17, 2025 | No shutdown date announced | gemini-3.5-flash | - | gemini-2.5-flash | June 17, 2025 | October 16, 2026 | gemini-3.5-flash | - | gemini-2.5-flash-preview-05-20 | May 20, 2025 | November 18, 2025 | gemini-3.5-flash | - | gemini-2.5-flash-preview-09-25 | September 25, 2025 | February 17, 2026 | gemini-3.5-flash | - | gemini-2.0-flash | February 5, 2025 | June 1, 2026 | gemini-3.5-flash | - | gemini-2.0-flash-001 | February 5, 2025 | June 1, 2026 | gemini-3.5-flash | + | gemini-3.6-flash | July 21, 2026 | No shutdown date announced | | + | gemini-3.5-flash-lite | July 21, 2026 | No shutdown date announced | | + | gemini-3.1-flash-lite | May 7, 2026 | May 7, 2027 | gemini-3.5-flash-lite | + | gemini-3-flash-preview | December 17, 2025 | No shutdown date announced | gemini-3.6-flash | + | gemini-2.5-flash | June 17, 2025 | October 16, 2026 | gemini-3.6-flash | + | gemini-2.5-flash-preview-05-20 | May 20, 2025 | November 18, 2025 | gemini-3.6-flash | + | gemini-2.5-flash-preview-09-25 | September 25, 2025 | February 17, 2026 | gemini-3.6-flash | + | gemini-2.0-flash | February 5, 2025 | June 1, 2026 | gemini-3.6-flash | + | gemini-2.0-flash-001 | February 5, 2025 | June 1, 2026 | gemini-3.6-flash |
- Pricing Jul 21, 2026, 6:00 PM
Google Gemini API pricing changed
Google Gemini API introduces tiered pricing across new Flash, Flash Lite, and Pro models with rates from $0.05 to $13.50 per 1M tokens
Gemini 2.0 Flash pricing defined: $1.50 input, $7.50 output per 1M tokensGemini 2.0 Flash Lite pricing defined: $0.75 input, $3.75 output per 1M tokensGemini 2.0 Pro pricing defined: $2.70 input, $13.50 output per 1M tokensGemini 3 Flash pricing defined: $0.30 input (text/image/video/audio), multiple output rates per 1M tokensGemini 3 Flash Lite pricing defined: $0.15 input, $1.25 output per 1M tokensGemini 3.1 Pro pricing defined: $0.54 input, $4.50 output per 1M tokensGemini 3.1 Flash pricing defined: $0.05 input, $0.20 output per 1M tokensView raw diff +10 −0
+ $7.50 + $3.75 + $3.75 + $13.50 + $0.30 (text / image / video / audio) + $0.03 + $0.15 (text / image / video / audio) + $0.15 (text / image / video / audio) + $0.54 (text / image / video / audio) + $0.05
- Update Jul 10, 2026, 6:00 AM
Google Gemini API changelog changed
Model codename changed from 'Nano Banana Lite' to 'Nano Banana 2 Lite' for gemini-3.1-flash-lite-image.
- gemini-3.1-flash-lite-image codename changed from 'Nano Banana Lite' to 'Nano Banana 2 Lite'
View raw diff +1 −1
- Released gemini-3.1-flash-lite-image (Nano Banana Lite) to general + Released gemini-3.1-flash-lite-image (Nano Banana 2 Lite) to general
- Update Jul 9, 2026, 6:00 PM
Google Gemini API changelog changed
Developer logs for Interactions API calls are now viewable in the AI Studio dashboard.
- Interactions API logs support added
- Logs viewable in AI Studio dashboard
View raw diff +2 −0
+ July 6, 2026 + Developer logs support for the Interactions API: logs for supported Interactions API calls are now viewable in the AI Studio dashboard.
- Update Jul 2, 2026, 12:00 PM
Google Gemini deprecations changed
Added deprecation timeline information for gemini-embedding-2 and embedding-2-preview models with specific dates.
- gemini-embedding-2 has April 22, 2026 deprecation date with no shutdown date announced
- embedding-2-preview deprecates March 10, 2026 and shuts down August 10, 2026
- embedding-2-preview will be replaced by gemini-embedding-2
View raw diff +2 −1
- Gemini deprecations + | gemini-embedding-2 | April 22, 2026 | No shutdown date announced | | + | embedding-2-preview | March 10, 2026 | August 10, 2026 | gemini-embedding-2 |
- Update Jul 1, 2026, 12:00 AM
Google Gemini API changelog changed
Two new Gemini models released: gemini-omni-flash-preview in public preview for video generation, and gemini-3.1-flash-lite-image moved to general availability.
- gemini-omni-flash-preview released in public preview for high-speed video generation (3-10 second videos at 720p)
- gemini-3.1-flash-lite-image (Nano Banana Lite) released to general availability for image generation and editing
View raw diff +13 −0
+ June 30, 2026 + Gemini Omni Flash in public preview: Released gemini-omni-flash-preview, + a high-performance multimodal model designed for high-speed video generation + and conversational video editing. Using the Interactions API, + you can generate 3--10 second videos at 720p from text descriptions or animate still images, + and then conversationally edit and refine the outputs. To get started, see the + Gemini Omni Flash guide and the + Gemini Omni Flash model card. + Released gemini-3.1-flash-lite-image (Nano Banana Lite) to general + availability (GA), our built-in multimodal model optimized for ultra-low + latency and cost-effective image generation and editing. See the [Gemini 3.1 + Flash Lite Image](https://ai.google.dev/gemini-api/docs/models/gemini-3.1-flash-lite-image) model + card and the Image generation guide.
- Pricing Jun 30, 2026, 6:00 PM
Google Gemini API pricing changed
Google Gemini API pricing updated with new rates for multiple models including text output at $1.50/$9.00/$0.25/$0.125 per 1M tokens and video output at $17.50 per 1M tokens with image pricing at $30 per 1M tokens.
now$1.50$9.00$17.505,792tokens$0.10$0.25$0.03$0.12View raw diff +11 −0
+ $1.50 (text / image / video / audio) + $9.00 (text) + $17.50 (video)* + Billing is based on total output token consumption, calculated at a rate of 5,792 tokens per second of 720p video. Under Standard pricing, this equates to an effective price of approximately $0.10 per second. + $0.25 (text/image/video) + Equivalent to $0.0336 per 1K resolution image* + $0.125 (text/image/video) + $0.75 (text and thinking) + $15.00 (images) + Equivalent to $0.0168 per 1K resolution image* + Image output is priced at $30 per 1,000,000 tokens. Output images at 1K (1024x1024px) consume 1120 tokens and are equivalent to $0.0336 per image.
- Update Jun 26, 2026, 6:00 AM
Google Gemini API changelog changed
Gemini 1.5 Flash is now the model behind gemini-flash-latest.
- gemini-flash-latest now points to Gemini 1.5 Flash
View raw diff +1 −1
- agentic and coding tasks. + agentic and coding tasks. This is now the model behind gemini-flash-latest.