G
Gemini 2.5 Flash Lite
Reviewed 2026-07-22Googlegoogle/gemini-2.5-flash-lite
- Context
- 1.05M
- Price
- $0.1/$0.4
- Tooling
- 3/5
- Deploy
- Cloud API
Best for
- Long documents (~1.05M context)
- Cost-sensitive, high-volume workloads
Skip if
- You need open weights to self-host
Capability scores
Tool calling3/5
Multi-step reasoning2/5
Specifications
- Context window
- 1.05M tokens
- VRAM
- N/A (cloud-hosted)
- License
- proprietary
- Pricing
- $0.1/M input · $0.4/M output
- Deployment
- Cloud API
- Docs
- Model docs ↗
Our take
Gemini 2.5 Flash Lite from Google offers a 1.05M-token context window and is well-suited to long-context work. Pricing is $0.1/M input · $0.4/M output, billed per token through the provider API.
Watch for the trade-offs: weights are closed, so you cannot self-host.
Compare to
Not sure this is the one?
Describe your project and let LMFinder rank the field for your exact constraints.
Describe your project →