Google AI (Gemini API) changed its pricing page
ChangePricing
Google Gemini API introduced new Gemini 3.8 Flash model with pricing across four tiers (Standard, Batch, Flex, Priority), launched Lyria 3.5 music generation model, and adjusted pricing effective dates and marketing descriptions for existing models.
Gemini 3.7 Flash is now available. Try it out.
Gemini 3.8 Flash is now available. Try it out.
GEMINI 3.8 FLASH
gemini-3.8-flash
Our most intelligent Flash model, engineered for long-horizon software engineering, autonomous agents, and complex enterprise workflows.
STANDARD
Free Tier Paid Tier, per 1M tokens in USD Input price Free of charge $0.75 through December 31, 2026. $1.50 starting January 1, 2027. Output price (including thinking tokens) Free of charge $3.75 through December 31, 2026. $7.50 starting January 1, 2027.
BATCH
Input price Not available $0.375 through December 31, 2026. $0.75 starting January 1, 2027. Output price (including thinking tokens) Not available $1.875 through December 31, 2026. $3.75 starting January 1, 2027.
FLEX
Input price Not available $0.375 through December 31, 2026. $0.75 starting January 1, 2027. Output price (including thinking tokens) Not available $1.875 through December 31, 2026. $3.75 starting January 1, 2027.
PRIORITY
Input price Free of charge $1.35 through December 31, 2026. $2.70 starting January 1, 2027. Output price (including thinking tokens) Free of charge $6.75 through December 31, 2026. $13.50 starting January 1, 2027.
Confidence95%
Full diff
===================================================================
--- before
+++ after
@@ -27,9 +27,9 @@
* 한국어
Get API key Cookbook Community Sign in
-Gemini 3.7 Flash is now available. Try it out.
+Gemini 3.8 Flash is now available. Try it out.
* Home
*
Gemini API
*
@@ -74,15 +74,59 @@
* check_circleML ops, model garden and more
Contact Sales
+GEMINI 3.8 FLASH
+
+gemini-3.8-flash
+
+Try it in Google AI Studio
+
+Our most intelligent Flash model, engineered for long-horizon software engineering, autonomous agents, and complex enterprise workflows.
+
+STANDARD
+
+Free Tier Paid Tier, per 1M tokens in USD Input price Free of charge $0.75 through December 31, 2026.
+$1.50 starting January 1, 2027. Output price (including thinking tokens) Free of charge $3.75 through December 31, 2026.
+$7.50 starting January 1, 2027. Context caching price Free of charge $0.075 through December 31, 2026.
+$0.15 starting January 1, 2027.
+$0.50 / 1,000,000 tokens per hour (storage price) through December 31, 2026.
+$1.00 / 1,000,000 tokens per hour (storage price) starting January 1, 2027. Grounding with Google Search* Not available 5,000 free search requests per month (shared across all Gemini 3.x models), then $14 per 1,000 requests. Grounding with Google Maps Not available 5,000 prompts per month (free, shared across Gemini 3), then $14 / 1,000 search queries Used to improve our products Yes No
+
+BATCH
+
+Free Tier Paid Tier, per 1M tokens in USD Input price Not available $0.375 through December 31, 2026.
+$0.75 starting January 1, 2027. Output price (including thinking tokens) Not available $1.875 through December 31, 2026.
+$3.75 starting January 1, 2027. Context caching price Not available $0.0375 through December 31, 2026.
+$0.075 starting January 1, 2027.
+$0.50 / 1,000,000 tokens per hour (storage price) through December 31, 2026.
+$1.00 / 1,000,000 tokens per hour (storage price) starting January 1, 2027. Used to improve our products Yes No
+
+FLEX
+
+Free Tier Paid Tier, per 1M tokens in USD Input price Not available $0.375 through December 31, 2026.
+$0.75 starting January 1, 2027. Output price (including thinking tokens) Not available $1.875 through December 31, 2026.
+$3.75 starting January 1, 2027. Context caching price Not available $0.0375 through December 31, 2026.
+$0.075 starting January 1, 2027.
+$0.50 / 1,000,000 tokens per hour (storage price) through December 31, 2026.
+$1.00 / 1,000,000 tokens per hour (storage price) starting January 1, 2027. Used to improve our products Yes No
+
+PRIORITY
+
+Free Tier Paid Tier, per 1M tokens in USD Input price Free of charge $1.35 through December 31, 2026.
+$2.70 starting January 1, 2027. Output price (including thinking tokens) Free of charge $6.75 through December 31, 2026.
+$13.50 starting January 1, 2027. Context caching price Free of charge $0.135 through December 31, 2026.
+$0.27 starting January 1, 2027.
+$0.50 / 1,000,000 tokens per hour (storage price) through December 31, 2026.
+$1.00 / 1,000,000 tokens per hour (storage price) starting January 1, 2027. Grounding with Google Search* Not available** 5,000 free search requests per month (shared across all Gemini 3.x models), then $14 per 1,000 requests. Grounding with Google Maps Not available** 5,000 prompts per month (free, shared across Gemini 3), then $14 / 1,000 search queries Used to improve our products Yes No
+
GEMINI 3.7 FLASH
gemini-3.7-flash
Try it in Google AI Studio
-Our most capable Flash model for agentic workflows and multimodal reasoning.
+Our high-speed, efficient Flash model built for everyday coding, agentic tool use, and reliable multi-step execution.
STANDARD
Free Tier Paid Tier, per 1M tokens in USD Input price Free of charge $0.75 through December 31, 2026.
@@ -124,9 +168,9 @@
gemini-3.6-flash
Try it in Google AI Studio
-A Flash model built for speed, combining frontier intelligence with superior search and grounding.
+Our previous generation Flash model, balancing speed and multimodal capabilities across general agentic and everyday tasks.
STANDARD
Free Tier Paid Tier, per 1M tokens in USD Input price Free of charge $0.75 through December 31, 2026.
@@ -172,9 +216,9 @@
gemini-3.5-flash
Try it in Google AI Studio
-A Flash model built for speed, combining frontier intelligence with superior search and grounding.
+Our earlier Flash model, built for speed and foundational performance across routine, high-throughput workloads.
STANDARD
Free Tier Paid Tier, per 1M tokens in USD Input price Free of charge $1.50 Output price (including thinking tokens) Free of charge $9.00 Context caching price Free of charge $0.15
@@ -481,9 +525,9 @@
gemini-3-flash-preview
Try it in Google AI Studio
-An intelligent model built for speed, combining intelligence with search and grounding.
+Our legacy Flash model, providing baseline speed and intelligence.
STANDARD
Free Tier Paid Tier, per 1M tokens in USD Input price Free of charge $0.50 (text / image / video)
@@ -759,13 +803,21 @@
(4k output not supported) Used to improve our products Yes No
Note: In some cases, an audio processing issue may prevent a video from being generated. You will only be charged if your video is successfully generated.
+LYRIA 3.5
+
+lyria-3.5
+
+Google's music generation model.
+
+Free Tier Paid Tier, per request in USD Lyria 3.5 (Full Song) Not available $0.08 per song Used to improve our products Yes No
+
LYRIA 3
lyria-3-clip-preview and lyria-3-pro-preview
-Google's family of music generation models.
+Google's family of legacy music generation models.
Free Tier Paid Tier, per request in USD Lyria 3 Clip Preview (30s) Not available $0.04 per song Lyria 3 Pro Preview (Full Song) Not available $0.08 per song Used to improve our products Yes No
GEMINI EMBEDDING 2
@@ -887,14 +939,15 @@
Model Tools Gemini Deep Research agent All model inference is charged at standard Gemini list rates, including input, output, and intermediate input / reasoning tokens generated during agentic loops. Tool usage fees apply per existing pricing structure, maintaining standard distinctions for Search Grounding (retrieved tokens excluded) versus Url_context / File Search (retrieved tokens included in all other tools). Managed agents in Gemini API All model inference is charged at standard Gemini list rates, including input, output, and intermediate input / reasoning tokens generated during agentic loops. (See pricing details). Environment compute (CPU, memory, sandbox execution) is not billed during the preview period. Antigravity Agent All model inference is charged at standard Gemini list rates, including input, output, and intermediate input / reasoning tokens generated during agentic loops. (See pricing details). Environment compute (CPU, memory, sandbox execution) is not billed during the preview period.
NOTES
+ * Agentic video understanding: When using agentic video understanding, token usage is variable based on the content loaded by the model rather than full video length. This typically results in up to 88% fewer input tokens for long-form video, though token counts depend on query complexity and dynamic sampling depth (which may exceed 1 FPS for detailed visual segments). See Agentic video understanding.
* Document token billing: Tokens for the DOCUMENT modality (for example, PDFs) are billed at the image token rate. In API responses, these tokens appear under the DOCUMENT modality within promptTokensDetails.
* Google AI Studio usage is free of charge in all available regions. See Billing FAQs for details.
* Prices may differ from the prices listed here and the prices offered on Gemini Enterprise Agent Platform. For Gemini Enterprise Agent Platform prices, see the Gemini Enterprise Agent Platform pricing page.
* If you are using dynamic retrieval to optimize costs, only requests that contain at least one grounding support URL from the web in their response are charged for Grounding with Google Search. Costs for Gemini always apply. Rate limits are subject to change.
Except as otherwise noted, the content of this page is licensed under the Creative Commons Attribution 4.0 License, and code samples are licensed under the Apache 2.0 License. For details, see the Google Developers Site Policies. Java is a registered trademark of Oracle and/or its affiliates.
-Last updated 2026-08-28 UTC.
+Last updated 2026-09-04 UTC.
-[[["Easy to understand","easyToUnderstand","thumb-up"],["Solved my problem","solvedMyProblem","thumb-up"],["Other","otherUp","thumb-up"]],[["Missing the information I need","missingTheInformationINeed","thumb-down"],["Too complicated / too many steps","tooComplicatedTooManySteps","thumb-down"],["Out of date","outOfDate","thumb-down"],["Samples / code issue","samplesCodeIssue","thumb-down"],["Other","otherDown","thumb-down"]],["Last updated 2026-08-28 UTC."],[],[]]
\ No newline at end of file
+[[["Easy to understand","easyToUnderstand","thumb-up"],["Solved my problem","solvedMyProblem","thumb-up"],["Other","otherUp","thumb-up"]],[["Missing the information I need","missingTheInformationINeed","thumb-down"],["Too complicated / too many steps","tooComplicatedTooManySteps","thumb-down"],["Out of date","outOfDate","thumb-down"],["Samples / code issue","samplesCodeIssue","thumb-down"],["Other","otherDown","thumb-down"]],["Last updated 2026-09-04 UTC."],[],[]]
\ No newline at end of file
Get the next one in your inbox.
Follow the vendor for free, or write your own prompt and watch any page the same way.