Analysis Summary
Gemini 2.5 Flash Lite is Google's most affordable multimodal model, supporting text, image, file, audio, and video inputs with a 1 million token context window. At $0.10 input and $0.40 output per million tokens, it is one of the cheapest options with genuine multimodal capability. Tool use and function calling are included. Its intelligence index of 6.9 and livecodebench of 0.400 reflect modest but functional performance.
For businesses, Flash Lite suits high-volume, cost-sensitive workflows: document classification, image tagging, audio transcription pipelines, and lightweight content generation. The large context window is a practical advantage for batch document processing. Reasoning depth and coding reliability are limited, so it is not suited for complex analysis or autonomous agent tasks.
Teams looking to process large volumes of mixed-media content at minimal cost will find Flash Lite a strong fit. Pair it with a more capable model for tasks requiring deeper reasoning.
Assessed July 10, 2026
Editorial notes
Gemini 2.5 Flash Lite from Google offers full multimodal input support and a 1M token context at just $0.10 input per million tokens, though reasoning and coding benchmarks are limited.
Rankings consider pricing, capabilities, benchmarks, and real-world applicability and are refreshed as new models launch. Feedback?
DFO Verdict
Gemini 2.5 Flash Lite from Google offers full multimodal input support and a 1M token context at just $0.10 input per million tokens, though reasoning and coding benchmarks are limited.
How Google: Gemini 2.5 Flash Lite compares
Google: Gemini 2.5 Flash Lite ranks #295 of 395 AI models we track for overall intelligence, #249 of 302 for agentic tasks. Its 1M-token context window is larger than 96% of the models we list. At $0.10 per million input tokens it is cheaper than 68% of comparable models.
Dark bar = input Ā· light bar = output, scaled to the priciest peer.
1M tokens ā 8 full-length novels or ~2,500 pages of business documents in a single request.
Strongest on value. The pulled-in intelligence corner is the trade-off, and if the shape matters more than the price, this is your model.
Compare shapes side-by-side āPricing
| Token Type | Cost per 1M tokens | Cost per 1K tokens |
|---|---|---|
| Input | $0.10 | $0.000100 |
| Output | $0.40 | $0.000400 |
What would Google: Gemini 2.5 Flash Lite cost your business?
Pick the job that looks most like yours, then fine-tune with the sliders. Estimates update live.
A website chatbot handling around 100 customer conversations a day, a few short messages each.
Full calculator with 620 models ā Price Calculator
These numbers get smaller with the right architecture.
We route routine calls to cheap models and save Google: Gemini 2.5 Flash Lite for the hard ones. Most clients cut their estimate by 60-80%.
Talk to our teamAbout Google: Gemini 2.5 Flash Lite
Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance..
Explore Related Models
Frequently asked questions about Google: Gemini 2.5 Flash Lite
How much does Google: Gemini 2.5 Flash Lite cost?
Google: Gemini 2.5 Flash Lite costs $0.10 per million input tokens and $0.40 per million output tokens.
What is the context window of Google: Gemini 2.5 Flash Lite?
Google: Gemini 2.5 Flash Lite has a context window of 1,048,576 tokens (1M).
What can Google: Gemini 2.5 Flash Lite do?
Google: Gemini 2.5 Flash Lite supports image/vision input, tool use, and function calling.
Who created Google: Gemini 2.5 Flash Lite?
Google: Gemini 2.5 Flash Lite is developed by Google and was released on July 22, 2025.
Data sourced from the OpenRouter API, Artificial Analysis, the Hugging Face Open LLM Leaderboard and our own internal testing. Scores are editorially curated by our team.
Last updated: July 25, 2026 8:38 pm