Gemini 2.5 Flash Lite
by Google
Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token ge
Models
Gemini 2.5 Flash Lite
Added 10 July 2026
Overview
Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...
Best for
Best for
Ultra-cheap bulk multimodal and extraction
Use cases
- inbox triage
- OCR pipelines
- cheap RAG generation
How to use / API access
Call via OpenRouter with model id `google/gemini-2.5-flash-lite`.
- openrouter ·
google/gemini-2.5-flash-liteDocs
OpenRouter id: google/gemini-2.5-flash-lite
Benchmarks
- Intelligence Index (artificial-analysis) : 11.4 index
Enterprise DNA angle
Use Flash-Lite as the Omni firehose worker; promote failures to Flash or Pro.
Notes
Gemini 2.5 Flash Lite is listed in the EDNA Models directory from the OpenRouter catalogue.
Pros
- Aggressive price/performance for bulk jobs
- Multimodal with large context
Cons
- Weaker on hard reasoning than Pro/Flash
- Closed model
Get the free Developer’s Field Guide
A 27-page field guide to the AI coding workflow with Claude. Claude Code, MCP servers, the prompt patterns that work, and what to delegate. Free.
Enter your work email. We send it straight over, plus a few short notes worth knowing. Unsubscribe any time.