Google: Gemini 3.1 flash lite preview

google/gemini-3.1-flash-lite-preview

Gemini 3.1 Flash Lite Preview is Google's high-efficiency model optimized for high-volume use cases. It outperforms Gemini 2.5 Flash Lite on overall quality and approaches Gemini 2.5 Flash performance across key capabilities. Improvements span audio input/ASR, RAG snippet ranking, translation, data extraction, and code completion. Supports full thinking levels (minimal, low, medium, high) for fine-grained cost/performance trade-offs. Priced at half the cost of Gemini 3 Flash.

Model specifications

Input
text, image, video, audio, file
Output
text
Context
1,048,576 tokens
Max output
65,535 tokens
Input price
$0.25 / 1M tokens
Output price
$1.5 / 1M tokens
Released
2026-03-04

Capabilities

  • Streaming
  • Function calling
  • Vision
  • JSON mode
  • Playground

Provider pricing, discounts and data privacy

Compare effective provider prices, published discounts, regions, retention policies, training use, compliance, and privacy links by service tier.

Standard service tier

2 available providers · discounts up to 25.000000000000007% · tier input average $0.21875 / 1M tokens · tier output average $1.3125 / 1M tokens

Google Vertex

Tier: Standard · Region: US

Pricing
Input
$0.1875 / 1M tokens
Output
$1.125 / 1M tokens
Cache read
$0.01875 / 1M tokens
Discounts
  • Input: 25% off — list $0.25 / 1M tokens, discounted $0.1875 / 1M tokens, effective $0.1875 / 1M tokens
  • Cache read: 25% off — list $0.025 / 1M tokens, discounted $0.01875 / 1M tokens, effective $0.01875 / 1M tokens
  • Output: 25% off — list $1.5 / 1M tokens, discounted $1.125 / 1M tokens, effective $1.125 / 1M tokens
Data privacy and compliance
Region
US
Zero data retention
Yes
Data retention
Zero retention
Used for training
No
Data collection
Moderated
No
GDPR compliant
Yes
HIPAA compliant
Yes
SOC 2 certified
Yes
BYOK supported
Yes

Google AI Studio

Tier: Standard · Region: US

Pricing
Input
$0.25 / 1M tokens
Output
$1.5 / 1M tokens
Cache read
$0.025 / 1M tokens

No provider discount is currently published.

Data privacy and compliance
Region
US
Zero data retention
No
Data retention
55-day retention
Used for training
No
Data collection
Moderated
No
GDPR compliant
Yes
HIPAA compliant
No
SOC 2 certified
No
BYOK supported
Yes

Frequently asked questions

What is Google: Gemini 3.1 flash lite preview?
Gemini 3.1 Flash Lite Preview is Google's high-efficiency model optimized for high-volume use cases. It outperforms Gemini 2.5 Flash Lite on overall quality and approaches Gemini 2.5 Flash performance across key capabilities. Improvements span audio input/ASR, RAG snippet ranking, translation, data extraction, and code completion. Supports full thinking levels (minimal, low, medium, high) for fine-grained cost/performance trade-offs. Priced at half the cost of Gemini 3 Flash.
How much does Google: Gemini 3.1 flash lite preview cost?
Input costs start at $0.25 / 1M tokens and output costs start at $1.5 / 1M tokens. Provider-level prices vary by service tier.
Are provider discounts available for Google: Gemini 3.1 flash lite preview?
Yes. Current provider offers include discounts of up to 25% from list price. The provider table shows list, discounted, and effective prices.
What is the context length of Google: Gemini 3.1 flash lite preview?
Google: Gemini 3.1 flash lite preview supports a 1,048,576 token context window and up to 65,535 output tokens.
What capabilities does Google: Gemini 3.1 flash lite preview support?
Google: Gemini 3.1 flash lite preview supports Streaming, Function calling, Vision, JSON mode, Playground.
Which providers offer Google: Gemini 3.1 flash lite preview?
Google: Gemini 3.1 flash lite preview is available from Google Vertex, Google AI Studio.
How do providers handle data privacy for Google: Gemini 3.1 flash lite preview?
2 of 2 providers report that customer data is not used for training, and 1 offer zero-data-retention routing. Retention, compliance, and privacy-policy links are listed per provider.

Browse all AI models