Guides / Language models

Google Gemini models explained: 3.1 Pro, 3.8 Flash and the rest

Google is moving fast: Gemini 3.6 Flash came out on July 21, 3.7 Flash on August 13 and 3.8 Flash on September 2. Here’s how to make sense of the lineup.

The lineup at a glance

Model Status Input / output per 1M tokens Cached input
Gemini 3.1 Pro Preview $2 / $12 (up to 200K), $4 / $18 above $0.20
Gemini 3.8 Flash Stable $0.75 / $3.75 until Dec 31, 2026 $0.075
Gemini 3.7 Flash Stable $0.75 / $3.75 until Dec 31, 2026 $0.075
Gemini 3.6 Flash Stable $0.75 / $3.75 until Dec 31, 2026 $0.075
Gemini 3.5 Flash Stable, older $1.50 / $9 $0.15
Gemini 3.5 Flash-Lite Stable $0.30 / $2.50 $0.03

Gemini 3.1 Pro and 3.8 Flash both have a context window of about 1 million tokens (1,048,576) and up to 65,536 output tokens.

Pick 3.8 Flash for almost everything

Google calls Gemini 3.8 Flash “our most intelligent workhorse model.” It costs the same as 3.6 and 3.7 Flash, so there is little reason to use the older ones unless you have tested and tuned a prompt for them.

Watch the promo end date. The $0.75 / $3.75 price for the 3.6, 3.7 and 3.8 Flash models runs through December 31, 2026. From January 1, 2027 the price doubles to $1.50 / $7.50. If you’re budgeting for next year, use the higher number.

Gemini 3.1 Pro

3.1 Pro is Google’s strongest model in the API, released February 19, 2026, and still labelled a preview. It is cheap for a top-tier model: $2 input and $12 output per million tokens for prompts up to 200K tokens. Longer prompts cost $4 / $18.

Flash-Lite for volume

Gemini 3.5 Flash-Lite, at $0.30 / $2.50, is Google’s budget option for classification, extraction and other high-volume tasks. Compare it against GPT-6 Luna and DeepSeek V4.1 Flash on the price list.

Caching and batch

Cached input on current Gemini models is listed at one tenth of the normal input price. Explicit caching also charges a storage fee per hour. The Batch API costs 50% less.

Google AI plans for consumers

Plan Price per month What stands out
Free $0 Gemini 3.6 Flash, some 3.1 Pro access, 15 GB storage
Google AI Plus $4.99 2x Free’s limits, 400 GB storage
Google AI Pro $19.99 4x Free’s limits, 5 TB storage, 3.8 Flash access
Google AI Ultra $99.99 (5x Pro’s limits) or $199.99 (20x Pro’s limits) From 20 TB storage, highest limits

Gemini 3.8 Flash is available in the Gemini app for Pro and Ultra subscribers. See all AI plans side by side on the plans page.

Quick recommendations

  • Default API choice: 3.8 Flash, while the promo price lasts.
  • Hardest tasks: 3.1 Pro, keeping prompts under 200K tokens to avoid the higher rate.
  • Cheapest bulk work: 3.5 Flash-Lite.

See how people rate Gemini today on the dumb meter.

Sources

  1. Gemini API pricing
  2. Gemini 3.1 Pro announcement
  3. Gemini 3.1 Pro model page
  4. Gemini 3.8 Flash announcement
  5. Gemini 3.8 Flash model page
  6. Gemini 3.7 Flash announcement
  7. Gemini 3.6 Flash and 3.5 Flash-Lite announcement
  8. Google AI plans