Google: Gemini 2.5 Flash Lite

Google: Gemini 2.5 Flash Lite

google · Released Jul 22, 2025
Intelligence #218 / 743
39.3 our score
Speed #4 / 326
374.0 tok/s
Input Price #226 / 743
$0.100 per 1M tokens
Output Price #265 / 743
$0.400 per 1M tokens
Context #42 / 743
1M tokens

Analysis Summary

Google Gemini 2.5 Flash Lite is a lightweight multimodal model designed for broad input handling rather than difficult reasoning. It accepts text, images, files, audio, and video, supports tool use and function calling, and provides a 1M token context window. Those capabilities give it an unusually wide operational range for a low-cost model.

It suits document extraction, media classification, basic SEO production, content transformation, and high-volume customer workflows with predictable instructions. The measured reasoning capability is limited, so complex analysis, original strategy, and high-stakes client-facing work should be routed elsewhere. There is no supplied coding or agentic benchmark evidence to support stronger claims in those areas.

Its pricing is the central advantage. Use it as a volume tier for multimodal processing and routine structured output, with a stronger model handling exceptions and difficult decisions.

Assessed August 9, 2026

Editorial notes

Gemini 2.5 Flash Lite offers a 1M token context, vision, audio, video, file handling, tool use, and function calling at very low pricing. Its limited reasoning capability makes it best for routine, high-volume automation.

Rankings consider pricing, capabilities, benchmarks, and real-world applicability and are refreshed as new models launch. Feedback?

DFO Verdict

Gemini 2.5 Flash Lite offers a 1M token context, vision, audio, video, file handling, tool use, and function calling at very low pricing. Its limited reasoning capability makes it best for routine, high-volume automation.

#218 of 743 overall Down 3 this week

How Google: Gemini 2.5 Flash Lite compares

Google: Gemini 2.5 Flash Lite ranks #260 of 432 AI models we track for overall intelligence. Its 1M-token context window is larger than 94% of the models we list. At $0.10 per million input tokens it is cheaper than 70% of comparable models.

Position in the field
Intelligence: smarter than 71% of models #218
Speed: faster than 99% of models #4
Price: cheaper than 70% of models #226
Context: larger than 94% of models #42
worst in fieldmedianbest in field
Price vs frontier peers · $ per 1M tokens
Google: Gemini 2.5 Flash Lite $0.10 in $0.40 out
Claude Opus 5 $5.00 in $25.00 out
OpenAI: GPT-5.6 Sol $2.00 in $10.00 out
SpaceXAI: Grok 4.6 $2.00 in $6.00 out

Dark bar = input · light bar = output, scaled to the priciest peer.

Context window vs peers · tokens
Google: Gemini 2.5 Flash Lite 1M

1M tokens ≈ 8 full-length novels or ~2,500 pages of business documents in a single request.

Intelligence1.6Technical0Value8.3Content5.4
Performance profile

Strongest on value. The pulled-in technical corner is the trade-off, and if the shape matters more than the price, this is your model.

Compare shapes side-by-side →

Pricing

Token Type Cost per 1M tokens Cost per 1K tokens
Input $0.10 $0.000100
Output $0.40 $0.000400

What would Google: Gemini 2.5 Flash Lite cost your business?

Pick the job that looks most like yours, then fine-tune with the sliders. Estimates update live.

A website chatbot handling around 100 customer conversations a day, a few short messages each.

3,000
One request is one message, email, draft or automation call.
1,200 tokens

$0/mo Google: Gemini 2.5 Flash Lite

Full calculator with 743 models → Price Calculator

DFO AI AUTOMATION

These numbers get smaller with the right architecture.

We route routine calls to cheap models and save Google: Gemini 2.5 Flash Lite for the hard ones. Most clients cut their estimate by 60-80%.

Talk to our team

About Google: Gemini 2.5 Flash Lite

Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance..

Frequently asked questions about Google: Gemini 2.5 Flash Lite

How much does Google: Gemini 2.5 Flash Lite cost?

Google: Gemini 2.5 Flash Lite costs $0.10 per million input tokens and $0.40 per million output tokens.

What is the context window of Google: Gemini 2.5 Flash Lite?

Google: Gemini 2.5 Flash Lite has a context window of 1,048,576 tokens (1M).

What can Google: Gemini 2.5 Flash Lite do?

Google: Gemini 2.5 Flash Lite supports image/vision input, tool use, and function calling.

Who created Google: Gemini 2.5 Flash Lite?

Google: Gemini 2.5 Flash Lite is developed by Google and was released on July 22, 2025.

© 2026 Design for Online Ltd. Registered in England and Wales No. 10328553. VAT Registered. Design for Online® and Forerunner® are registered trademarks.