# Google: Gemini 2.5 Flash Lite (batch)

Provider: google  
Category: llm  
Model ID: `google/gemini-2.5-flash-lite:batch`

Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...

## Specs

- Context length: 1048576 tokens
- Max output: 65535 tokens
- Modalities: text, image, file, audio, video
- Released: 2025-07-22

## Pricing

- Input: $0.05 per million tokens
- Output: $0.20 per million tokens

## Verdict

Gemini 2.5 Flash Lite (batch) trades real-time responsiveness for extreme cost efficiency—$0.05/Mtok input makes it the cheapest multimodal model in Google's lineup. The batch-only constraint means jobs queue for up to 24 hours, ruling out interactive use but unlocking massive-scale processing of documents, images, audio, and video. Reach for this when you're processing thousands of files overnight and cost per token matters more than latency.

---
Last verified: 2026-09-04T02:00:56.012Z  
Canonical URL: https://switchy.build/directory/models/gemini-2-5-flash-lite-batch