Google: Gemini 2.5 Flash Lite

Google: Gemini 2.5 Flash Lite

google · Released Jul 22, 2025
Intelligence #162 / 620
45.4 our score
Speed #17 / 289
245.6 tok/s
Input Price #197 / 620
$0.100 per 1M tokens
Output Price #229 / 620
$0.400 per 1M tokens
Context #26 / 620
1M tokens

Analysis Summary

Gemini 2.5 Flash Lite is Google's most affordable multimodal model, supporting text, image, file, audio, and video inputs with a 1 million token context window. At $0.10 input and $0.40 output per million tokens, it is one of the cheapest options with genuine multimodal capability. Tool use and function calling are included. Its intelligence index of 6.9 and livecodebench of 0.400 reflect modest but functional performance.

For businesses, Flash Lite suits high-volume, cost-sensitive workflows: document classification, image tagging, audio transcription pipelines, and lightweight content generation. The large context window is a practical advantage for batch document processing. Reasoning depth and coding reliability are limited, so it is not suited for complex analysis or autonomous agent tasks.

Teams looking to process large volumes of mixed-media content at minimal cost will find Flash Lite a strong fit. Pair it with a more capable model for tasks requiring deeper reasoning.

Assessed July 10, 2026

Editorial notes

Gemini 2.5 Flash Lite from Google offers full multimodal input support and a 1M token context at just $0.10 input per million tokens, though reasoning and coding benchmarks are limited.

Rankings consider pricing, capabilities, benchmarks, and real-world applicability and are refreshed as new models launch. Feedback?

DFO Verdict

Gemini 2.5 Flash Lite from Google offers full multimodal input support and a 1M token context at just $0.10 input per million tokens, though reasoning and coding benchmarks are limited.

#162 of 620 overall

How Google: Gemini 2.5 Flash Lite compares

Google: Gemini 2.5 Flash Lite ranks #295 of 395 AI models we track for overall intelligence, #249 of 302 for agentic tasks. Its 1M-token context window is larger than 96% of the models we list. At $0.10 per million input tokens it is cheaper than 68% of comparable models.

Position in the field
Intelligence: smarter than 74% of models #162
Speed: faster than 94% of models #17
Price: cheaper than 68% of models #197
Context: larger than 96% of models #26
worst in fieldmedianbest in field
Price vs frontier peers Ā· $ per 1M tokens
Google: Gemini 2.5 Flash Lite $0.10 in $0.40 out
Anthropic: Claude Fable 5 $10.00 in $50.00 out
Anthropic: Claude Opus 4.8 $5.00 in $25.00 out
Claude Opus 5 $5.00 in $25.00 out

Dark bar = input Ā· light bar = output, scaled to the priciest peer.

Context window vs peers Ā· tokens
Google: Gemini 2.5 Flash Lite 1M

1M tokens ā‰ˆ 8 full-length novels or ~2,500 pages of business documents in a single request.

Intelligence1Technical1.5Value8.3Content5.5
Performance profile

Strongest on value. The pulled-in intelligence corner is the trade-off, and if the shape matters more than the price, this is your model.

Compare shapes side-by-side →

Pricing

Token Type Cost per 1M tokens Cost per 1K tokens
Input $0.10 $0.000100
Output $0.40 $0.000400

What would Google: Gemini 2.5 Flash Lite cost your business?

Pick the job that looks most like yours, then fine-tune with the sliders. Estimates update live.

A website chatbot handling around 100 customer conversations a day, a few short messages each.

3,000
One request is one message, email, draft or automation call.
1,200 tokens

$0/mo Google: Gemini 2.5 Flash Lite

Full calculator with 620 models → Price Calculator

DFO AI AUTOMATION

These numbers get smaller with the right architecture.

We route routine calls to cheap models and save Google: Gemini 2.5 Flash Lite for the hard ones. Most clients cut their estimate by 60-80%.

Talk to our team

About Google: Gemini 2.5 Flash Lite

Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance..

Frequently asked questions about Google: Gemini 2.5 Flash Lite

How much does Google: Gemini 2.5 Flash Lite cost?

Google: Gemini 2.5 Flash Lite costs $0.10 per million input tokens and $0.40 per million output tokens.

What is the context window of Google: Gemini 2.5 Flash Lite?

Google: Gemini 2.5 Flash Lite has a context window of 1,048,576 tokens (1M).

What can Google: Gemini 2.5 Flash Lite do?

Google: Gemini 2.5 Flash Lite supports image/vision input, tool use, and function calling.

Who created Google: Gemini 2.5 Flash Lite?

Google: Gemini 2.5 Flash Lite is developed by Google and was released on July 22, 2025.