Gemini 2.5 Flash
Gemini 2.5 Flash is an AI model from Google, available on ZeroTwo. It costs $0.30 per million input tokens and $2.50 per million output tokens, has a 1,000,000-token context window, can return up to 65,536 output tokens, and has a knowledge cutoff of January 2025.
A Gemini 2.5 model balancing cost and capability for large scale processing, low-latency, high volume tasks that require thinking, and agentic use cases. Accepts text, image, video, and audio input and returns text, with a 1M token context window and adjustable thinking budgets.
The $0.30 input and $0.03 cached input rates cover text, image and video. Audio input is billed separately at $1.00 per million tokens, and cached audio input at $0.10 per million tokens.
Questions
How much does Gemini 2.5 Flash cost?
Gemini 2.5 Flash is priced at $0.30 per million input tokens and $2.50 per million output tokens, with cached input at $0.03 per million. On ZeroTwo it is included in your plan's credits rather than billed per token.
What is Gemini 2.5 Flash's context window?
Gemini 2.5 Flash accepts up to 1,000,000 tokens of context and can return up to 65,536 output tokens.
What is Gemini 2.5 Flash's knowledge cutoff date?
Gemini 2.5 Flash has a training knowledge cutoff of January 2025. For anything later than that, use a model with web search enabled.
Can I use Gemini 2.5 Flash for free?
You can try Gemini 2.5 Flash on ZeroTwo's free plan, which includes a monthly credit allowance across every model in the catalogue. Heavier use moves to a paid plan rather than a per-token bill.