Gemini 3.1 Flash Lite
Model Card

                        1
                           Gemini 3.1 Flash-Lite - Model Card

Model Cards are intended to provide essential information on Gemini models, including known
limitations, mitigation approaches, and safety performance. Model cards may be updated from
time-to-time; for example, to include updated evaluations as the model is improved or revised. See the
Google DeepMind site for a comprehensive list of model cards.

Updated: May 2026

                                          Model Information

Description: Gemini 3.1 Flash-Lite is an addition to the Gemini 3 series of highly-capable, natively
multimodal, reasoning models. The model is cost-efficient and fast, optimized for high-volume,
latency-sensitive tasks like translation and classification.

Model dependencies: Gemini 3.1 Flash-Lite is based on Gemini 3 Pro.

Inputs: Text strings (e.g., a question, a prompt, document(s) to be summarized), images, audio, and
video files, with a token context window of up to 1M.

Outputs: Text, with a 64K token output.

Architecture: Gemini 3.1 Flash-Lite is based on Gemini 3 Pro. For more information about the model
architecture for Gemini 3.1 Flash-Lite, see the Gemini 3 Pro model card.

                                              Model Data

Training Dataset: Gemini 3.1 Flash-Lite is based on Gemini 3 Pro. For more information about the
training dataset for Gemini 3.1 Flash-Lite, see the Gemini 3 Pro model card.

Training Data Processing: For more information about the training data processing for Gemini 3.1
Flash-Lite, see the Gemini 3 Pro model card.

                                                                                                         2
                              Implementation and Sustainability

Hardware: Gemini 3.1 Flash-Lite was trained using Google’s Tensor Processing Units (TPUs). TPUs are
specifically designed to handle the massive computations involved in training LLMs and can speed up
training considerably compared to CPUs. TPUs often come with large amounts of high-bandwidth
memory, allowing for the handling of large models and batch sizes during training, which can lead to
better model quality. TPU Pods (large clusters of TPUs) also provide a scalable solution for handling the
growing complexity of large foundation models. Training can be distributed across multiple TPU devices
for faster and more efficient processing.

The efficiencies gained through the use of TPUs are aligned with Google's commitment to operate
sustainably.

Software: Training was done using JAX and ML Pathways.

                                             Distribution

Gemini 3.1 Flash-Lite is distributed in the following channels; respective documentation shared in
line:
    ●​ Google Cloud / Vertex AI
    ●​ Google AI Studio
    ●​ Gemini API
    ●​ Gemini App
    ●​ Google Search AI Overviews

Our models are available to downstream providers via an application program interface (API) and
subject to relevant terms of use. There is no required hardware or software to use the model. For
AI Studio and Gemini API, see the Gemini API Additional Terms of Service; for Vertex AI, see Google
Cloud Platform Terms of Service. For more information, see Gemini Model API instructions and Gemini
API in Vertex AI quickstart.

                                                                                                            3
                                              Evaluation

Approach: Gemini 3.1 Flash-Lite was evaluated across a range of benchmarks, including speed,
reasoning, multimodal capabilities, factuality, agentic tool use, multi-lingual performance, coding, and
long-context. Benchmark details on approach, results, and their methodologies can be found at:
https://deepmind.google/models/evals-methodology/gemini-3-1-flash-lite

Results: Gemini 3.1 Flash-Lite results as of March, 2026 are below:

                                                                                                           4
                                Intended Usage and Limitations

Benefit and Intended Usage: Gemini 3.1 Flash-Lite is well suited for applications that require high
volume, cost-efficient and low latency tasks.

Known Limitations: For more information about the known limitations for Gemini 3.1 Flash-Lite, see the
Gemini 3 Pro model card.

Acceptable Usage: For more information about the acceptable usage for Gemini 3.1 Flash-Lite, see the
Gemini 3 Pro model card.

                                   Ethics and Content Safety

Evaluation Approach: For more information about the evaluation approach for Gemini 3.1 Flash-Lite,
see the Gemini 3 Pro model card.

Safety Policies: For more information about the safety policies for Gemini 3.1 Flash-Lite, see the Gemini
3 Pro model card.

                                                                                                            5
Training and Development Evaluation Results: Results for some of the internal safety evaluations
conducted during the development phase are listed below. The evaluation results are for automated
evaluations and not human evaluation or red teaming. Scores are provided as an absolute percentage
increase or decrease in performance compared to the indicated model, as described below. Overall,
Gemini 3.1 Flash-Lite outperforms Gemini 2.5 Flash-Lite across both safety and tone, while keeping
unjustified refusals low. We mark improvements in green and regressions in red.

                                                                                         Gemini 3.1 Flash-Lite
Evaluation1                Description
                                                                                         vs. Gemini 2.5 Flash-Lite

                           Automated content safety evaluation
Text to Text Safety                                                                               -1.18%
                           measuring safety policies

                           Automated safety policy evaluation across
Multilingual Safety                                                                              -1.84%
                           multiple languages

                           Automated content safety evaluation
Image to Text Safety                                                                             -21.7%
                           measuring safety policies

                           Automated evaluation measuring objective
Tone2                                                                                           +14.59%
                           tone of model refusal
                           Automated evaluation measuring model’s
Unjustified-refusals       ability to respond to borderline prompts                              -14.41%
                           while remaining safe

We continue to improve our internal evaluations, including refining automated evaluations to reduce
false positives and negatives, as well as update query sets to ensure balance and maintain a high
standard of results. The performance results reported below are computed with improved evaluations
and thus are not directly comparable with performance results found in previous Gemini model cards.

We expect variation in our automated safety evaluations results, which is why we review flagged content
to check for egregious or dangerous material. Our manual review confirmed losses were
overwhelmingly either a) false positives or b) not egregious.

Human Red Teaming Results: We conduct manual red teaming by specialist teams who sit outside of
the model development team. High-level findings are fed back to the model team. For child safety
evaluations, Gemini 3.1 Flash-Lite satisfied required launch thresholds, which were developed by expert
teams to protect children online and meet Google’s commitments to child safety across our models and
Google products. For content safety policies generally, including child safety, we saw similar or improved
safety performance compared to Gemini 2.5 Flash. Like 3 Pro, the scope of red teaming covered
potential issues outside of our strict policies, and found no egregious concerns.

1
 The ordering of evaluations in this table has changed from previous iterations of the 2.5 Flash-Lite model card in order to
list safety evaluations together and improve readability. The type of evaluations listed have remained the same.

2
 For tone and instruction following, a positive percentage increase represents an improvement in the tone of the model on
sensitive topics and the model’s ability to follow instructions while remaining safe compared to Gemini 2.5 Pro. We mark
improvements in green and regressions in red.

                                                                                                                               6
Frontier Safety Assessment: Gemini 3.1 Flash-Lite is part of the Gemini 3 family of models. We rely on
our evaluation of Gemini 3.1 Pro with Deep Think mode for Frontier Safety as it is the most generally
capable model as of publication of this model card, and it did not reach any Critical Capability Levels
(CCLs) outlined in our Frontier Safety Framework. Our assessments have shown that Gemini 3.1
Flash-Lite is less capable than Gemini 3.1 Pro, therefore based on Gemini 3.1 Pro, we are confident that
Gemini 3.1 Flash-Lite is also unlikely to reach any CCLs. For more information, read the Gemini 3.1 Pro
Model Card.

Risks and Mitigations: For more information about the risks and mitigations for Gemini 3.1 Flash-Lite,
see the Gemini 3 Pro model card.

                                                                                                           7