Google AI (Gemini API)
ai.google.dev · AI · watching since Aug 17, 2026 · RSS
Current state
As of Sep 6, 2026Google AI's Gemini API offers a Free tier with limited model access and free tokens, a Paid tier billed per-model at per-1M-token rates (with Standard, Batch, Flex, and Priority pricing options), and a quote-only Enterprise tier via Gemini Enterprise Agent Platform. Pricing for flagship models like Gemini 3.8 Flash includes promotional rates through December 31, 2026, rising January 1, 2027; additional models cover embeddings, TTS, video (Veo), music (Lyria), robotics, and tool usage (Search, Maps, code execution).
| Plan | Price | Includes |
|---|---|---|
| Free | $0 | Limited access to certain models · Free input & output tokens · Google AI Studio access · Content used to improve our products* |
| Paid | — | Higher rate limits for production deployments · Access to Context caching · Batch API (50% cost reduction) · Access to Google's most advanced models · Content not used to improve our products* |
| Enterprise | — | All features in Paid, plus optional access to: · Dedicated support channels · Advanced security & compliance · Provisioned throughput · Volume-based discounts (based on usage) |
| Gemini 3.8 Flash (Standard) | $0.75 per 1M tokens (input) | Output $3.75 per 1M tokens (through Dec 31, 2026) · Prices rise to $1.50 input / $7.50 output starting Jan 1, 2027 · Context caching $0.075 (through Dec 31, 2026), $0.15 after · Grounding with Google Search: 5,000 free requests/month, then $14/1,000 · Content not used to improve products (paid tier) |
| Gemini 3.8 Flash (Batch) | $0.375 per 1M tokens (input) | Output $1.875 per 1M tokens (through Dec 31, 2026) · Prices rise to $0.75 input / $3.75 output starting Jan 1, 2027 · Context caching $0.0375 (through Dec 31, 2026), $0.075 after |
| Gemini 3.8 Flash (Flex) | $0.375 per 1M tokens (input) | Output $1.875 per 1M tokens (through Dec 31, 2026) · Prices rise to $0.75 input / $3.75 output starting Jan 1, 2027 · Context caching $0.0375 (through Dec 31, 2026), $0.075 after |
| Gemini 3.8 Flash (Priority) | $1.35 per 1M tokens (input) | Output $6.75 per 1M tokens (through Dec 31, 2026) · Prices rise to $2.70 input / $13.50 output starting Jan 1, 2027 · Context caching $0.135 (through Dec 31, 2026), $0.27 after · Free tier available (free of charge) |
| Gemini 3.5 Flash (Standard) | $1.50 per 1M tokens (input) | Output $9.00 per 1M tokens · Context caching $0.15, storage $1.00/1M tokens/hour · Grounding with Google Search: 5,000 free requests/month, then $14/1,000 |
| Gemini Embedding 2 (Standard) | $0.20 per 1M tokens (text input) | Image input $0.45 ($0.00012 per image) · Audio input $6.50 ($0.00016 per second) · Video input $12.00 ($0.00079 per frame) |
| Gemini Embedding (Standard) | $0.15 per 1M tokens | Batch price $0.075 per 1M tokens · Free of charge on Free Tier |
| Veo 3.1 | $0.40 per second (720p/1080p) | $0.60 per second for 4k · Veo 3.1 Fast: $0.10 (720p), $0.12 (1080p), $0.30 (4k) · Veo 3.1 Lite: $0.05 (720p), $0.08 (1080p), no 4k · Charged only if video successfully generated |
| Lyria 3.5 | $0.08 per song | Full song generation · Not available on Free Tier |
Free tier Free tier provides limited access to certain models, free input & output tokens, Google AI Studio access; content used to improve products. Google AI Studio usage itself is free of charge in all available regions.
Changes
- ChangePricing
Google Gemini API introduced new Gemini 3.8 Flash model with pricing across four tiers (Standard, Batch, Flex, Priority), launched Lyria 3.5 music generation model, and adjusted pricing effective dates and marketing descriptions for existing models.
Gemini 3.7 Flash is now available. Try it out.Gemini 3.8 Flash is now available. Try it out. GEMINI 3.8 FLASH gemini-3.8-flash Our most intelligent Flash model, engineered for long-horizon software engineering, autonomous agents, and complex enterprise workflows. STANDARD Free Tier Paid Tier, per 1M tokens in USD Input price Free of charge $0.75 through December 31, 2026. $1.50 starting January 1, 2027. Output price (including thinking tokens) Free of charge $3.75 through December 31, 2026. $7.50 starting January 1, 2027. BATCH Input price Not available $0.375 through December 31, 2026. $0.75 starting January 1, 2027. Output price (including thinking tokens) Not available $1.875 through December 31, 2026. $3.75 starting January 1, 2027. FLEX Input price Not available $0.375 through December 31, 2026. $0.75 starting January 1, 2027. Output price (including thinking tokens) Not available $1.875 through December 31, 2026. $3.75 starting January 1, 2027. PRIORITY Input price Free of charge $1.35 through December 31, 2026. $2.70 starting January 1, 2027. Output price (including thinking tokens) Free of charge $6.75 through December 31, 2026. $13.50 starting January 1, 2027.Confidence95% - ChangePricing
Gemini Omni Flash model moved from preview to generally available on paid tier with pricing: $1.50 per 1M input tokens and $9.00 per 1M text output tokens / $17.50 per 1M video output tokens.
GEMINI OMNI FLASH PREVIEW gemini-omni-flash-previewGEMINI OMNI FLASH gemini-omni-1.1-flash Try it in Google AI Studio Our next-generation video generation and editing model, now generally available to developers on the paid tier of the Gemini API. STANDARD Free Tier Paid Tier, per 1M tokens in USD Input price Not available $1.50 (text / image / video / audio) Output price (including thinking tokens) Not available $9.00 (text) $17.50 (video)* Used to improve our products Yes No * Billing is based on total output token consumption, calculated at a rate of 5,792 tokens per second of 720p video. Under Standard pricing, this equates to an effective price of approximately $0.10 per second. GEMINI OMNI FLASH PREVIEW gemini-omni-flash-previewConfidence95% - ChangePricing
Google Gemini API pricing page: Added two new transcription models (Gemini 3.5 Transcribe Live and Gemini 3.5 Transcribe) with pricing, removed deprecated Imagen 4 and older Veo models, removed 'Preview' rate limit disclaimers, and updated various model descriptions.
The Interactions API is now generally available. We recommend using this API for access to all the latest features and models.Gemini 3.7 Flash is now available. Try it out.Confidence85%
This is the Competitor pricing pack running on a real vendor.
Current state
As of Sep 6, 2026This changelog tracks Gemini API releases, including new models (e.g., Gemini 3.8 Flash, Lyria 3.5, Gemini 3.5 Transcribe), feature launches (agentic video understanding, streaming TTS, Computer Use), and deprecation notices for older model versions. Entries span from December 2023 through September 2026, with the most recent update dated September 3, 2026.
Deprecations
- Sunset 2026-09-30
gemini-omni-flash-preview
→ gemini-omni-1.1-flash
- Sunset 2026-08-31
gemini-robotics-er-1.6-preview
- No date
temperature, top_p, top_k sampling parameters
- Sunset 2026-08-17
imagen-4.0-generate-001, imagen-4.0-ultra-generate-001, imagen-4.0-fast-generate-001
- Sunset 2026-06-30
veo-2.0-generate-001, veo-3.0-generate-001, veo-3.0-fast-generate-001
→ veo-3.1-generate-preview, veo-3.1-fast-generate-preview
- Sunset 2026-06-15
GMP Contextual View tool
- Sunset 2026-06-25
gemini-3.1-flash-image-preview, gemini-3-pro-image-preview
→ gemini-3.1-flash-image, gemini-3-pro-image
- Sunset 2026-05-25
gemini-3.1-flash-lite-preview
→ gemini-3.1-flash-lite
- Sunset 2026-03-31
gemini-2.5-flash-lite-preview-09-2025
→ gemini-3.1-flash-lite-preview
- Sunset 2026-04-30
gemini-robotics-er-1.5-preview
→ gemini-robotics-er-1.6-preview
- Sunset 2026-03-09
gemini-3-pro-preview
→ gemini-3.1-pro-preview
- Sunset 2026-06-08
Interactions API legacy schema (outputs field)
→ new steps schema
Latest entries
- Lyria 3.5 released in public preview for full-length music generationaddition
- Gemini 3.8 Flash reaches general availabilityaddition
- Agentic video understanding released for Gemini 3.7 Flash, 3.6 Flash, and 3.5 Flash-Liteaddition
- Gemini Omni Flash reaches general availability with video extension and interpolationaddition
- Gemini 3.5 Transcribe and Transcribe Live reach general availabilityaddition
- Gemini 3.7 Flash reaches general availabilityaddition
- Gemini Robotics ER 2 released in public previewaddition
- Gemini-robotics-er-1.6-preview will be shut down on August 31, 2026deprecation
- Gemini 3.6 Flash and Gemini 3.5 Flash-Lite reach general availabilityaddition
- Sampling parameters temperature, top_p and top_k are now deprecateddeprecation
- Developer logs support added for Interactions API callsaddition
- Gemini Omni Flash released in public previewaddition
Changes
- ChangeAPI changelog
Gemini 3.8 Flash released as GA; Lyria 3.5 and agentic video understanding now in preview; three Imagen and Veo models deprecated with shutdown dates.
Gemini 3.7 Flash is now available. Try it out.Gemini 3.8 Flash is now available. Try it out. SEPTEMBER 3, 2026 * Lyria 3.5 in public preview: Released the next generation of Google's music generation model: lyria-3.5: Full-length song generation... SEPTEMBER 2, 2026 * Gemini 3.8 Flash generally available (GA): Released gemini-3.8-flash... SEPTEMBER 1, 2026 * Agentic video understanding: Released agentic video understanding for Gemini 3.7 Flash, 3.6 Flash, and 3.5 Flash-Lite...Confidence95% - ChangeAPI changelog
Gemini Omni Flash GA released with new video capabilities; gemini-omni-flash-preview endpoint deprecated on September 30, 2026.
AUGUST 26, 2026 * Gemini 3.5 Transcribe generally available (GA): Released two dedicated speech-to-text models based on Gemini's audio understanding:AUGUST 27, 2026 * Gemini Omni Flash generally available (GA): Released gemini-omni-1.1-flash, the GA version of our fast, conversational video generation and editing model. This release includes significant new capabilities: * Video extension: Seamlessly extend existing videos by generating continuations at the end of a clip using the extend task or directly with a prompt. * Interpolation (first + last frame): Generate a video transitioning between two images using the image_to_video task with up to 2 images. * Resolution control: New resolution parameter in video_config supports 360p, 720p (default), 1080p, and 4k outputs. 1080p and 4K outputs are generated using upscaling. The existing gemini-omni-flash-preview endpoint will be deprecated on September 30, 2026. To get started, see the Gemini Omni Flash model page and the omni guide. AUGUST 26, 2026 * Gemini 3.5 Transcribe generally available (GA): Released two dedicated speech-to-text models based on Gemini's audio understanding:Confidence95% - ChangeAPI changelog
Gemini 3.5 Transcribe models released as generally available (GA) on August 26, 2026, with new speech-to-text capabilities including streaming support via Live API.
AUGUST 13, 2026 * Gemini 3.7 Flash generally available (GA): Released our most intelligent workhorse model yet for coding and agents:AUGUST 26, 2026 * Gemini 3.5 Transcribe generally available (GA): Released two dedicated speech-to-text models based on Gemini's audio understanding: * Gemini 3.5 Transcribe (gemini-3.5-transcribe): High-accuracy, low-latency non-streaming speech-to-text with utterance-based language detection across 85+ languages, speaker diarization, word-level timestamps, and custom vocabulary biasing (up to 1,000 terms). * Gemini 3.5 Transcribe Live (gemini-3.5-transcribe-live): Low-latency, bidirectional streaming speech-to-text over WebSockets using the Live API, supporting interim and finalized transcription events, Smart transcription mode, and multiple Voice Activity Detection (VAD) strategies. To get started, see the Audio transcription guide, the Live transcription guide, and the Gemini 3.5 Transcribe model page. AUGUST 13, 2026 * Gemini 3.7 Flash generally available (GA): Released our most intelligent workhorse model yet for coding and agents:Confidence95% - ChangeAPI changelog
Semantic Retriever removed from Gemini API beta features list.
* New beta features: * Function Calling * Semantic Retriever * Attributed Question Answering (AQA)* New beta features: * Function Calling * Attributed Question Answering (AQA)Confidence95%
This is the Product and API pack running on a real vendor.
Watch Google AI (Gemini API) — or anything else — your way.
Your own prompt, cadence and channels. One alert with the before/after proof when your sentence comes true.