Gemini 3.1 Flash-Lite
Gemini 3.1 Flash-Lite is an AI model from Google, available on ZeroTwo. It costs $0.25 per million input tokens and $1.50 per million output tokens, has a 1,048,576-token context window, can return up to 65,536 output tokens, and has a knowledge cutoff of January 2025.
A low-cost multimodal model in the Gemini 3.1 family, aimed at high-frequency, lightweight tasks. Google positions it for high-volume agentic work, simple data extraction, translation, transcription, and latency-sensitive applications where budget and speed are the primary constraints. Supports thinking for enhanced accuracy.
The $0.25 input and $0.025 cached input rates cover text, image and video. Audio input is billed separately at $0.50 per million tokens, and cached audio input at $0.05 per million tokens.
Questions
How much does Gemini 3.1 Flash-Lite cost?
Gemini 3.1 Flash-Lite is priced at $0.25 per million input tokens and $1.50 per million output tokens, with cached input at $0.025 per million. On ZeroTwo it is included in your plan's credits rather than billed per token.
What is Gemini 3.1 Flash-Lite's context window?
Gemini 3.1 Flash-Lite accepts up to 1,048,576 tokens of context and can return up to 65,536 output tokens.
What is Gemini 3.1 Flash-Lite's knowledge cutoff date?
Gemini 3.1 Flash-Lite has a training knowledge cutoff of January 2025. For anything later than that, use a model with web search enabled.
Can I use Gemini 3.1 Flash-Lite for free?
You can try Gemini 3.1 Flash-Lite on ZeroTwo's free plan, which includes a monthly credit allowance across every model in the catalogue. Heavier use moves to a paid plan rather than a per-token bill.