Skip to content
modsignal

Google AI (Gemini API)

ai.google.dev · AI · watching since Aug 17, 2026 · RSS

Follow
https://ai.google.dev/pricingChecked every 168 h · last checked 21 h ago · watching since Aug 17, 2026
Watch it yourself →

Current state

As of Sep 6, 2026

Google AI's Gemini API offers a Free tier with limited model access and free tokens, a Paid tier billed per-model at per-1M-token rates (with Standard, Batch, Flex, and Priority pricing options), and a quote-only Enterprise tier via Gemini Enterprise Agent Platform. Pricing for flagship models like Gemini 3.8 Flash includes promotional rates through December 31, 2026, rising January 1, 2027; additional models cover embeddings, TTS, video (Veo), music (Lyria), robotics, and tool usage (Search, Maps, code execution).

PlanPriceIncludes
Free$0Limited access to certain models · Free input & output tokens · Google AI Studio access · Content used to improve our products*
PaidHigher rate limits for production deployments · Access to Context caching · Batch API (50% cost reduction) · Access to Google's most advanced models · Content not used to improve our products*
EnterpriseAll features in Paid, plus optional access to: · Dedicated support channels · Advanced security & compliance · Provisioned throughput · Volume-based discounts (based on usage)
Gemini 3.8 Flash (Standard)$0.75 per 1M tokens (input)Output $3.75 per 1M tokens (through Dec 31, 2026) · Prices rise to $1.50 input / $7.50 output starting Jan 1, 2027 · Context caching $0.075 (through Dec 31, 2026), $0.15 after · Grounding with Google Search: 5,000 free requests/month, then $14/1,000 · Content not used to improve products (paid tier)
Gemini 3.8 Flash (Batch)$0.375 per 1M tokens (input)Output $1.875 per 1M tokens (through Dec 31, 2026) · Prices rise to $0.75 input / $3.75 output starting Jan 1, 2027 · Context caching $0.0375 (through Dec 31, 2026), $0.075 after
Gemini 3.8 Flash (Flex)$0.375 per 1M tokens (input)Output $1.875 per 1M tokens (through Dec 31, 2026) · Prices rise to $0.75 input / $3.75 output starting Jan 1, 2027 · Context caching $0.0375 (through Dec 31, 2026), $0.075 after
Gemini 3.8 Flash (Priority)$1.35 per 1M tokens (input)Output $6.75 per 1M tokens (through Dec 31, 2026) · Prices rise to $2.70 input / $13.50 output starting Jan 1, 2027 · Context caching $0.135 (through Dec 31, 2026), $0.27 after · Free tier available (free of charge)
Gemini 3.5 Flash (Standard)$1.50 per 1M tokens (input)Output $9.00 per 1M tokens · Context caching $0.15, storage $1.00/1M tokens/hour · Grounding with Google Search: 5,000 free requests/month, then $14/1,000
Gemini Embedding 2 (Standard)$0.20 per 1M tokens (text input)Image input $0.45 ($0.00012 per image) · Audio input $6.50 ($0.00016 per second) · Video input $12.00 ($0.00079 per frame)
Gemini Embedding (Standard)$0.15 per 1M tokensBatch price $0.075 per 1M tokens · Free of charge on Free Tier
Veo 3.1$0.40 per second (720p/1080p)$0.60 per second for 4k · Veo 3.1 Fast: $0.10 (720p), $0.12 (1080p), $0.30 (4k) · Veo 3.1 Lite: $0.05 (720p), $0.08 (1080p), no 4k · Charged only if video successfully generated
Lyria 3.5$0.08 per songFull song generation · Not available on Free Tier

Free tier Free tier provides limited access to certain models, free input & output tokens, Google AI Studio access; content used to improve products. Google AI Studio usage itself is free of charge in all available regions.

Changes

  1. ChangePricing

    Google Gemini API introduced new Gemini 3.8 Flash model with pricing across four tiers (Standard, Batch, Flex, Priority), launched Lyria 3.5 music generation model, and adjusted pricing effective dates and marketing descriptions for existing models.

    Gemini 3.7 Flash is now available. Try it out.
    Gemini 3.8 Flash is now available. Try it out. GEMINI 3.8 FLASH gemini-3.8-flash Our most intelligent Flash model, engineered for long-horizon software engineering, autonomous agents, and complex enterprise workflows. STANDARD Free Tier Paid Tier, per 1M tokens in USD Input price Free of charge $0.75 through December 31, 2026. $1.50 starting January 1, 2027. Output price (including thinking tokens) Free of charge $3.75 through December 31, 2026. $7.50 starting January 1, 2027. BATCH Input price Not available $0.375 through December 31, 2026. $0.75 starting January 1, 2027. Output price (including thinking tokens) Not available $1.875 through December 31, 2026. $3.75 starting January 1, 2027. FLEX Input price Not available $0.375 through December 31, 2026. $0.75 starting January 1, 2027. Output price (including thinking tokens) Not available $1.875 through December 31, 2026. $3.75 starting January 1, 2027. PRIORITY Input price Free of charge $1.35 through December 31, 2026. $2.70 starting January 1, 2027. Output price (including thinking tokens) Free of charge $6.75 through December 31, 2026. $13.50 starting January 1, 2027.
    Confidence95%
  2. ChangePricing

    Gemini Omni Flash model moved from preview to generally available on paid tier with pricing: $1.50 per 1M input tokens and $9.00 per 1M text output tokens / $17.50 per 1M video output tokens.

    GEMINI OMNI FLASH PREVIEW gemini-omni-flash-preview
    GEMINI OMNI FLASH gemini-omni-1.1-flash Try it in Google AI Studio Our next-generation video generation and editing model, now generally available to developers on the paid tier of the Gemini API. STANDARD Free Tier Paid Tier, per 1M tokens in USD Input price Not available $1.50 (text / image / video / audio) Output price (including thinking tokens) Not available $9.00 (text) $17.50 (video)* Used to improve our products Yes No * Billing is based on total output token consumption, calculated at a rate of 5,792 tokens per second of 720p video. Under Standard pricing, this equates to an effective price of approximately $0.10 per second. GEMINI OMNI FLASH PREVIEW gemini-omni-flash-preview
    Confidence95%
  3. ChangePricing

    Google Gemini API pricing page: Added two new transcription models (Gemini 3.5 Transcribe Live and Gemini 3.5 Transcribe) with pricing, removed deprecated Imagen 4 and older Veo models, removed 'Preview' rate limit disclaimers, and updated various model descriptions.

    The Interactions API is now generally available. We recommend using this API for access to all the latest features and models.
    Gemini 3.7 Flash is now available. Try it out.
    Confidence85%

This is the Competitor pricing pack running on a real vendor.

Watch Google AI (Gemini API) — or anything else — your way.

Your own prompt, cadence and channels. One alert with the before/after proof when your sentence comes true.

Get started