Skip to content

Google · Documents, media and automation

Gemini 3.8 Flash

Processes text, images, audio and video; the output is text.

Compare with another model Suitable combinations ↓

Costs
$0.75 input / $3.75 output per million tokens through December 31, 2026. From January 1, 2027: $1.50 / $7.50.
Speed
We have not independently measured speed on a comparable basis.
Availability
Gemini API and Google AI Studio
Input length
1,048,576 input tokens; output capped at 65,536 tokens.
Inputs
Text, images, video, audio and PDF → text

Data checked . Published September 2, 2026. Specifications and prices are provided by the vendor; usage recommendations come from the editorial team. We do not yet have our own comparative test of this version.

Ideal use

When to choose it

  • Analysis of documents, audio and video
  • Programming and multi-step business tasks

Usage boundaries

When to choose another model

  • A different model is needed to generate images or audio directly.

Traceable supporting sources

Data and measurement sources

Distinguish between the manufacturer's documentation and the results of a specific test. Measurements also depend on the settings and the task set used.

Google

Vendor documentation

Specifications, availability, and terms

Manufacturer data verified as of the review date. Recommended use is an editorial interpretation.

Open original source ↗
Google

Pricing

Price and its validity period

Standard API rates; the app subscription is billed separately.

Open original source ↗

Efficiency in practice

With what Gemini 3.8 Flash combine

Many tasks without unnecessary costs

A low-cost model handles routine work, and the more expensive one gets only the truly difficult cases.

Use your own data to verify when a stronger model should take over the work. A model cannot reliably assess its own answer.

Trends over time

Related events from AI Radar

Search more →
OpenAI important update

OpenAI opens the Pro plan to new users, but halves API credits and moves toward payment for actual usage

New: API credits per dollar halved; Pro plan at $200/month reopened to new users; 5-hour window for the weekly allotment removed; Strategic shift from subsidized subscription models to pay-per-use pricing; Microsoft is making similar changes with Copilot

Users considering the OpenAI Pro plan at $200/month now also benefit from the removal of the 5-hour limit, but those paying through the API receive fewer credits for the same dollar than before — the new lower prices of the GPT-6 Sol and Luna models only partially offset this.

Companies using the OpenAI API should recalculate their AI operating costs, because the number of credits per dollar is changing and the price of the new GPT-6 Sol and Luna models has also fallen by 50 % compared with the 5.6 series — the net impact on the budget depends on the specific usage volume and types of models used.

✓ 3
GitHub important update

GitHub recommends Claude Opus 5.5 following the retirement of Claude Opus 4.7

New: The recommended replacement for Claude Opus 4.7 is Claude Opus 5.5, not Claude Opus 5.

When switching from Claude Opus 4.7, follow the current recommendation to use Claude Opus 5.5; the earlier announcement listed a different replacement.

Migrating company workflows and integrations to supported models may require enabling replacement models through access policies in Copilot Enterprise or Copilot Business.

✓ official

Next step

Compare the model using your actual task.

A benchmark narrows the selection. A short trial on your data determines the choice.

Open comparison →