Google: Gemini 3.1 Flash Lite Preview

google/gemini-3.1-flash-lite-preview

Description

Gemini 3.1 Flash Lite Preview is Google's high-efficiency model optimized for high-volume use cases. It outperforms Gemini 2.5 Flash Lite on overall quality and approaches Gemini 2.5 Flash performance across key capabilities. Improvements span audio input/ASR, RAG snippet ranking, translation, data extraction, and code completion. Supports full thinking levels (minimal, low, medium, high) for fine-grained cost/performance trade-offs.

How this model compares

Overall covers the full catalog. By plan covers only models available on that tier (same rules as available models in your list). Position on min–average–max. Prices use the higher of prompt or completion per token, shown per 1M tokens.

Price (per 1M tokens)

Min

Max

This model

336 models in this groupPrice (per 1M tokens)

Min: $0.04
Avg: $12.570002
Max: $750.00

This model: $1.50 / 1M tokens

Context length (tokens)

Min

Max

This model

336 models in this groupContext length (tokens)

Min: 4,095 tokens
Avg: 398,336.839 tokens
Max: 2,000,000 tokens

This model: 1,048,576 tokens

Capabilities

text+image+file+audio+video->textContext: 1,048,576 tokens

Input:

ImageFileAudioVideoText

Output:

Text