Skip to content
modsignal

Google AI (Gemini API) changed its pricing page

Pricinghttps://ai.google.dev/pricing

ChangePricing

Google Gemini API introduced new Gemini 3.8 Flash model with pricing across four tiers (Standard, Batch, Flex, Priority), launched Lyria 3.5 music generation model, and adjusted pricing effective dates and marketing descriptions for existing models.

Gemini 3.7 Flash is now available. Try it out.
Gemini 3.8 Flash is now available. Try it out. GEMINI 3.8 FLASH gemini-3.8-flash Our most intelligent Flash model, engineered for long-horizon software engineering, autonomous agents, and complex enterprise workflows. STANDARD Free Tier Paid Tier, per 1M tokens in USD Input price Free of charge $0.75 through December 31, 2026. $1.50 starting January 1, 2027. Output price (including thinking tokens) Free of charge $3.75 through December 31, 2026. $7.50 starting January 1, 2027. BATCH Input price Not available $0.375 through December 31, 2026. $0.75 starting January 1, 2027. Output price (including thinking tokens) Not available $1.875 through December 31, 2026. $3.75 starting January 1, 2027. FLEX Input price Not available $0.375 through December 31, 2026. $0.75 starting January 1, 2027. Output price (including thinking tokens) Not available $1.875 through December 31, 2026. $3.75 starting January 1, 2027. PRIORITY Input price Free of charge $1.35 through December 31, 2026. $2.70 starting January 1, 2027. Output price (including thinking tokens) Free of charge $6.75 through December 31, 2026. $13.50 starting January 1, 2027.
Confidence95%
Full diff
===================================================================
--- before
+++ after
@@ -27,9 +27,9 @@
  * 한국어
 
 Get API key Cookbook Community Sign in
 
-Gemini 3.7 Flash is now available. Try it out.
+Gemini 3.8 Flash is now available. Try it out.
  * Home
  * 
    Gemini API
  * 
@@ -74,15 +74,59 @@
  * check_circleML ops, model garden and more
 
 Contact Sales
 
+GEMINI 3.8 FLASH
+
+gemini-3.8-flash
+
+Try it in Google AI Studio
+
+Our most intelligent Flash model, engineered for long-horizon software engineering, autonomous agents, and complex enterprise workflows.
+
+STANDARD
+
+Free Tier Paid Tier, per 1M tokens in USD Input price Free of charge $0.75 through December 31, 2026.
+$1.50 starting January 1, 2027. Output price (including thinking tokens) Free of charge $3.75 through December 31, 2026.
+$7.50 starting January 1, 2027. Context caching price Free of charge $0.075 through December 31, 2026.
+$0.15 starting January 1, 2027.
+$0.50 / 1,000,000 tokens per hour (storage price) through December 31, 2026.
+$1.00 / 1,000,000 tokens per hour (storage price) starting January 1, 2027. Grounding with Google Search* Not available 5,000 free search requests per month (shared across all Gemini 3.x models), then $14 per 1,000 requests. Grounding with Google Maps Not available 5,000 prompts per month (free, shared across Gemini 3), then $14 / 1,000 search queries Used to improve our products Yes No
+
+BATCH
+
+Free Tier Paid Tier, per 1M tokens in USD Input price Not available $0.375 through December 31, 2026.
+$0.75 starting January 1, 2027. Output price (including thinking tokens) Not available $1.875 through December 31, 2026.
+$3.75 starting January 1, 2027. Context caching price Not available $0.0375 through December 31, 2026.
+$0.075 starting January 1, 2027.
+$0.50 / 1,000,000 tokens per hour (storage price) through December 31, 2026.
+$1.00 / 1,000,000 tokens per hour (storage price) starting January 1, 2027. Used to improve our products Yes No
+
+FLEX
+
+Free Tier Paid Tier, per 1M tokens in USD Input price Not available $0.375 through December 31, 2026.
+$0.75 starting January 1, 2027. Output price (including thinking tokens) Not available $1.875 through December 31, 2026.
+$3.75 starting January 1, 2027. Context caching price Not available $0.0375 through December 31, 2026.
+$0.075 starting January 1, 2027.
+$0.50 / 1,000,000 tokens per hour (storage price) through December 31, 2026.
+$1.00 / 1,000,000 tokens per hour (storage price) starting January 1, 2027. Used to improve our products Yes No
+
+PRIORITY
+
+Free Tier Paid Tier, per 1M tokens in USD Input price Free of charge $1.35 through December 31, 2026.
+$2.70 starting January 1, 2027. Output price (including thinking tokens) Free of charge $6.75 through December 31, 2026.
+$13.50 starting January 1, 2027. Context caching price Free of charge $0.135 through December 31, 2026.
+$0.27 starting January 1, 2027.
+$0.50 / 1,000,000 tokens per hour (storage price) through December 31, 2026.
+$1.00 / 1,000,000 tokens per hour (storage price) starting January 1, 2027. Grounding with Google Search* Not available** 5,000 free search requests per month (shared across all Gemini 3.x models), then $14 per 1,000 requests. Grounding with Google Maps Not available** 5,000 prompts per month (free, shared across Gemini 3), then $14 / 1,000 search queries Used to improve our products Yes No
+
 GEMINI 3.7 FLASH
 
 gemini-3.7-flash
 
 Try it in Google AI Studio
 
-Our most capable Flash model for agentic workflows and multimodal reasoning.
+Our high-speed, efficient Flash model built for everyday coding, agentic tool use, and reliable multi-step execution.
 
 STANDARD
 
 Free Tier Paid Tier, per 1M tokens in USD Input price Free of charge $0.75 through December 31, 2026.
@@ -124,9 +168,9 @@
 gemini-3.6-flash
 
 Try it in Google AI Studio
 
-A Flash model built for speed, combining frontier intelligence with superior search and grounding.
+Our previous generation Flash model, balancing speed and multimodal capabilities across general agentic and everyday tasks.
 
 STANDARD
 
 Free Tier Paid Tier, per 1M tokens in USD Input price Free of charge $0.75 through December 31, 2026.
@@ -172,9 +216,9 @@
 gemini-3.5-flash
 
 Try it in Google AI Studio
 
-A Flash model built for speed, combining frontier intelligence with superior search and grounding.
+Our earlier Flash model, built for speed and foundational performance across routine, high-throughput workloads.
 
 STANDARD
 
 Free Tier Paid Tier, per 1M tokens in USD Input price Free of charge $1.50 Output price (including thinking tokens) Free of charge $9.00 Context caching price Free of charge $0.15
@@ -481,9 +525,9 @@
 gemini-3-flash-preview
 
 Try it in Google AI Studio
 
-An intelligent model built for speed, combining intelligence with search and grounding.
+Our legacy Flash model, providing baseline speed and intelligence.
 
 STANDARD
 
 Free Tier Paid Tier, per 1M tokens in USD Input price Free of charge $0.50 (text / image / video)
@@ -759,13 +803,21 @@
 (4k output not supported) Used to improve our products Yes No
 
 Note: In some cases, an audio processing issue may prevent a video from being generated. You will only be charged if your video is successfully generated.
 
+LYRIA 3.5
+
+lyria-3.5
+
+Google's music generation model.
+
+Free Tier Paid Tier, per request in USD Lyria 3.5 (Full Song) Not available $0.08 per song Used to improve our products Yes No
+
 LYRIA 3
 
 lyria-3-clip-preview and lyria-3-pro-preview
 
-Google's family of music generation models.
+Google's family of legacy music generation models.
 
 Free Tier Paid Tier, per request in USD Lyria 3 Clip Preview (30s) Not available $0.04 per song Lyria 3 Pro Preview (Full Song) Not available $0.08 per song Used to improve our products Yes No
 
 GEMINI EMBEDDING 2
@@ -887,14 +939,15 @@
 Model Tools Gemini Deep Research agent All model inference is charged at standard Gemini list rates, including input, output, and intermediate input / reasoning tokens generated during agentic loops. Tool usage fees apply per existing pricing structure, maintaining standard distinctions for Search Grounding (retrieved tokens excluded) versus Url_context / File Search (retrieved tokens included in all other tools). Managed agents in Gemini API All model inference is charged at standard Gemini list rates, including input, output, and intermediate input / reasoning tokens generated during agentic loops. (See pricing details). Environment compute (CPU, memory, sandbox execution) is not billed during the preview period. Antigravity Agent All model inference is charged at standard Gemini list rates, including input, output, and intermediate input / reasoning tokens generated during agentic loops. (See pricing details). Environment compute (CPU, memory, sandbox execution) is not billed during the preview period.
 
 NOTES
 
+ * Agentic video understanding: When using agentic video understanding, token usage is variable based on the content loaded by the model rather than full video length. This typically results in up to 88% fewer input tokens for long-form video, though token counts depend on query complexity and dynamic sampling depth (which may exceed 1 FPS for detailed visual segments). See Agentic video understanding.
  * Document token billing: Tokens for the DOCUMENT modality (for example, PDFs) are billed at the image token rate. In API responses, these tokens appear under the DOCUMENT modality within promptTokensDetails.
  * Google AI Studio usage is free of charge in all available regions. See Billing FAQs for details.
  * Prices may differ from the prices listed here and the prices offered on Gemini Enterprise Agent Platform. For Gemini Enterprise Agent Platform prices, see the Gemini Enterprise Agent Platform pricing page.
  * If you are using dynamic retrieval to optimize costs, only requests that contain at least one grounding support URL from the web in their response are charged for Grounding with Google Search. Costs for Gemini always apply. Rate limits are subject to change.
 
 Except as otherwise noted, the content of this page is licensed under the Creative Commons Attribution 4.0 License, and code samples are licensed under the Apache 2.0 License. For details, see the Google Developers Site Policies. Java is a registered trademark of Oracle and/or its affiliates.
 
-Last updated 2026-08-28 UTC.
+Last updated 2026-09-04 UTC.
 
-[[["Easy to understand","easyToUnderstand","thumb-up"],["Solved my problem","solvedMyProblem","thumb-up"],["Other","otherUp","thumb-up"]],[["Missing the information I need","missingTheInformationINeed","thumb-down"],["Too complicated / too many steps","tooComplicatedTooManySteps","thumb-down"],["Out of date","outOfDate","thumb-down"],["Samples / code issue","samplesCodeIssue","thumb-down"],["Other","otherDown","thumb-down"]],["Last updated 2026-08-28 UTC."],[],[]]
\ No newline at end of file
+[[["Easy to understand","easyToUnderstand","thumb-up"],["Solved my problem","solvedMyProblem","thumb-up"],["Other","otherUp","thumb-up"]],[["Missing the information I need","missingTheInformationINeed","thumb-down"],["Too complicated / too many steps","tooComplicatedTooManySteps","thumb-down"],["Out of date","outOfDate","thumb-down"],["Samples / code issue","samplesCodeIssue","thumb-down"],["Other","otherDown","thumb-down"]],["Last updated 2026-09-04 UTC."],[],[]]
\ No newline at end of file

Get the next one in your inbox.

Follow the vendor for free, or write your own prompt and watch any page the same way.

Get started