Skip to content

Gemini 3.8 Flash

Google's Flash model for coding and agents, with reasoning at Flash-level latency.

By Googlesource

Facts read from official pages, last on

Official page

Price

What the gateway charges per million tokens: what you send in, and what the model writes back.

Source: Vercel AI Gateway, checked 2026-09-25.

Input
$0.75per 1M tokens
Output
$3.75per 1M tokensOutput costs 5x the input.

Specs

How much it reads at once, how much it writes back, and how recent its knowledge is.

Source: Vercel AI Gateway, checked 2026-09-25.

Context window
1M tokens1,000,000
Max output
66K tokens65,535
Knowledge cutoff
Not listed
Released
Type
Language
Modalities
Takes text, image, PDF, video. Returns text.
Gateway ID
google/gemini-3.8-flash

Capabilities

What the gateway says this model supports.

Source: Vercel AI Gateway, checked 2026-09-25.

  • Reasoning
  • File input
  • Vision
  • Tool use
  • Web search
  • Implicit caching
  • Video input
  • Structured output

Data handling

What happens to your prompts, and what you have to switch on.

Source: Vercel AI Gateway, checked 2026-09-25.

ZDRSome providers

Zero data retention

Available on a paid Vercel plan through the AI Gateway, from some providers on it (you turn it on).

NTEvery provider

No training on prompts

Offered by every provider on the Vercel AI Gateway, when you turn it on.

Usage

How much people actually run it, as counted where it is served.

requests
189.8M189,838,598 requests
tokens
9.1T9,054,814,145,589 tokens

Counted on OpenRouter over the 24 days to 2026-09-25.

Also by Google15

Friday

The brief, by email

Five to seven links an issue, each with what to do about it.

Unsubscribe from any issue.

This page as Markdown