Gemini 3.1 Flash-Lite

google/gemini-3.1-flash-lite

Google Gemini 3.1 Flash-Lite high-volume model.

InputAudioFileImageTextVideo
OutputText
Input / Outputclub$0.265 / $1.59per 1M tokens · USD
Context window1Mtoken
Maximum output65.5Ktoken

Pricing

Current model usage rates. Choose your plan to see its published prices.

Pricing
UsagePricing
InputUS$0.265/ 1M token
Cache readUS$0.0265/ 1M token
Cache writeUS$0.0883333/ 1M token
Cache write · 1 hUS$0/ 1M token
OutputUS$1.59/ 1M token

Membership and model usage are billed separately. Compare plans

Capabilities

Input
AudioFileImageTextVideo
Output
Text
Context window
1,048,576 token
Maximum output
65,536 token

Providers

Model supply available through Vecbase.

Via OpenRouter

Access is subject to your account’s permissions and model availability.

Frequently asked questions

What is Gemini 3.1 Flash-Lite, and who develops it?

Gemini 3.1 Flash-Lite is a model from Google. Its model ID on Vecbase is google/gemini-3.1-flash-lite. Google Gemini 3.1 Flash-Lite high-volume model.

How much does Gemini 3.1 Flash-Lite cost on Vecbase?

For Gemini 3.1 Flash-Lite, the Club plan lists USD 0.265 per 1 million input tokens and USD 1.59 per 1 million output tokens, as of 9/15/2026, 4:36:00 AM UTC. See the pricing table for cache rates, input-length tiers and other usage types. Membership fees are separate.

How many tokens fit in Gemini 3.1 Flash-Lite’s context window?

Gemini 3.1 Flash-Lite has a context window of 1,048,576 tokens. Input and output must fit within the limits of the API you use.

What is the maximum output length for Gemini 3.1 Flash-Lite?

Gemini 3.1 Flash-Lite can generate up to 65,536 tokens per response. Your API or request settings may set a lower limit.

Can Gemini 3.1 Flash-Lite call tools in an agent workflow?

Yes. Gemini 3.1 Flash-Lite supports tool calling. Your application provides the tool definitions and runs the tools the model requests.

Does Gemini 3.1 Flash-Lite support structured outputs?

Yes. Gemini 3.1 Flash-Lite supports structured output. Set the required format in your API request and check the returned data before using it.

Which input and output formats does Gemini 3.1 Flash-Lite support?

Gemini 3.1 Flash-Lite accepts Audio, File, Image, Text, and Video and produces Text.

Which providers make Gemini 3.1 Flash-Lite available through Vecbase?

Vecbase offers Gemini 3.1 Flash-Lite through OpenRouter. Availability may vary by provider.

Which model ID should an agent use to request Gemini 3.1 Flash-Lite?

Use google/gemini-3.1-flash-lite as the model identifier for Gemini 3.1 Flash-Lite. See Vecbase’s API documentation for authentication, required inputs and supported request formats.

What does an agent need to access Gemini 3.1 Flash-Lite through Vecbase?

To use Gemini 3.1 Flash-Lite, you need an authenticated account with the required permissions, sufficient balance and the relevant capability enabled. See the documentation for setup.

Is a release date available for Gemini 3.1 Flash-Lite?

We do not yet have a confirmed release date for Gemini 3.1 Flash-Lite.

Which models can I compare with Gemini 3.1 Flash-Lite on Vecbase?

You can compare Gemini 3.1 Flash-Lite with Gemini Flash Latest (google/gemini-flash-latest), Gemini Pro Latest (google/gemini-pro-latest), and Vecbase 1.0 Lite (vecbase-1.0-lite). Check their capabilities and prices to find the best fit for your task.