Gemini 2.5 Flash Lite
Gemini 2.5 Flash Lite is an AI model from Google, available on ZeroTwo. It costs $0.10 per million input tokens and $0.40 per million output tokens, has a 1,000,000-token context window, can return up to 65,536 output tokens, and has a knowledge cutoff of January 2025.
Gemini 2.5 Flash Lite is a faster, cost-efficient version of Gemini 2.5 Flash. It's optimized for quick responses and cost-effective applications. Supports multimodal inputs including text, images, video, and audio.
The $0.10 input and $0.01 cached input rates cover text, image and video. Audio input is billed separately at $0.30 per million tokens, and cached audio input at $0.03 per million tokens.
Questions
How much does Gemini 2.5 Flash Lite cost?
Gemini 2.5 Flash Lite is priced at $0.10 per million input tokens and $0.40 per million output tokens, with cached input at $0.01 per million. On ZeroTwo it is included in your plan's credits rather than billed per token.
What is Gemini 2.5 Flash Lite's context window?
Gemini 2.5 Flash Lite accepts up to 1,000,000 tokens of context and can return up to 65,536 output tokens.
What is Gemini 2.5 Flash Lite's knowledge cutoff date?
Gemini 2.5 Flash Lite has a training knowledge cutoff of January 2025. For anything later than that, use a model with web search enabled.
Can I use Gemini 2.5 Flash Lite for free?
You can try Gemini 2.5 Flash Lite on ZeroTwo's free plan, which includes a monthly credit allowance across every model in the catalogue. Heavier use moves to a paid plan rather than a per-token bill.