Model comparison
Same task, different trade-offs.
We are not looking for an absolute winner. The comparison shows which model better suits your work, budget and required speed.
Models being compared
Gemini 3.5 Flash-Lite
Older profile · current status unverified
Data from July 31, 2026
Release date unverified
A model for quick classification, data retrieval and conversion into the required format. It is insufficient for complex decisions; its advantages are speed and low cost.
Gemini 3.8 Flash
Data from September 27, 2026
Published September 2, 2026
Processes text, images, audio and video; the output is text.
Price, speed and usage
Low · $0.30 per million input tokens and $2.50 per million output tokens; a token is a small piece of text
$0.75 input / $3.75 output per million tokens through December 31, 2026. From January 1, 2027: $1.50 / $7.50.
Very fast · generated 350 tokens per second in the measurement
We have not independently measured speed on a comparable basis.
Online service from Google and Gemini products
Gemini API and Google AI Studio
Very long input · up to 1 million tokens (small pieces of text)
1,048,576 input tokens; output capped at 65,536 tokens.
Text, images, video and audio as input
Text, images, video, audio and PDF → text
Gemini 3.5 Flash-Lite
Use when…
- Classifying requests and routing them to the appropriate workflow
- Finding data and then checking it
- Bulk rewrites and format conversions
Choose another model when…
- Autonomous solving of complex analytical tasks
- Final check of important claims
Gemini 3.8 Flash
Use when…
- Analysis of documents, audio and video
- Programming and multi-step business tasks
Choose another model when…
- A different model is needed to generate images or audio directly.
Original sources