Identity
- Status
- active
- Developer ID
gemini-3.1-flash-lite- Released
- Not reported
- Version
- Gemini 3.1 Flash-Lite
- License
- Unknown
Gemini 3.1 Flash-Lite is Google DeepMind's stable cost-efficient model for high-volume agentic tasks.
| Provider | Input / 1M | Output / 1M | Cache read / 1M | Cache write / 1M | Observed | Evidence |
|---|---|---|---|---|---|---|
Google AIAvailable | $0.25 | $1.50 | $0.025 | — | Aug 29, 2026 | Provider-reported |
| Gemini 3.1 Flash-LiteGoogle DeepMind | ToneBench2026-08-14-10-task | 69.50 | 1,571 | 50 | Gemini 3.1 Flash-Liteaverage per case | Original sourcewinner eligible | Third-party benchmark |
At a Glance
gemini-3.1-flash-liteThis page represents one immutable developer model ID. Serving endpoints, pricing, regions, quantizations, and service tiers attach separately and never create duplicate model records.
Primary Evidence
Comparable Peers
| A | Pair | B | Context |
|---|---|---|---|
Gemini 3.1 Flash-LiteGoogle DeepMind | vs | Gemini 2.5 Flash-LiteGoogle DeepMind | family variantstext |
Gemini 3.1 Flash-LiteGoogle DeepMind | vs | Gemini 3 FlashGoogle DeepMind | family variantstext |
Gemini 3.1 Flash-LiteGoogle DeepMind | vs | Nova 2 LiteAmazon | cross-developer peerstext |
Gemini 3.1 Flash-LiteGoogle DeepMind | vs | Qwen3.8-FlashQwen | cross-developer peerstext |
Gemini 3.1 Flash-LiteGoogle DeepMind | vs | Gemini Computer UseGoogle DeepMind | family variantstext |
Gemini 3.1 Flash-LiteGoogle DeepMind | vs | GPT-5.6 LunaOpenAI | cross-developer peerstext |
Gemini 3.1 Flash-LiteGoogle DeepMind | vs | Claude Opus 5Anthropic | cross-developer peerstext |
Gemini 3.1 Flash-LiteGoogle DeepMind | vs | Pixtral LargeMistral AI | cross-developer peerstext |
Gemini 3.1 Flash-LiteGoogle DeepMind | vs | GPT-4.1 MiniOpenAI | cross-developer peerstext |
Gemini 3.1 Flash-LiteGoogle DeepMind | vs | DeepSeek-V4-FlashDeepSeek | cross-developer peerstext |
Gemini 3.1 Flash-LiteGoogle DeepMind | vs | Sonar Reasoning ProPerplexity | cross-developer peerstext |
Gemini 3.1 Flash-LiteGoogle DeepMind | vs | Sonar ProPerplexity | cross-developer peerstext |
Questions
Gemini 3.1 Flash-Lite is Google DeepMind's stable cost-efficient model for high-volume agentic tasks. It is developed by Google DeepMind and its current sourced lifecycle status is active.
Gemini 3.1 Flash-Lite's current record lists Text, Image, Video, Audio, and Document as input and Text as output.
The current record lists a 1,048,576-token context window and a maximum output of 65,536 tokens.
Gemini 3.1 Flash-Lite is marked as API-available. The directly sourced provider records currently include Google AI.
The lowest directly sourced prices currently attached to Gemini 3.1 Flash-Lite are $0.25 per million input tokens through Google AI and $1.50 per million output tokens through Google AI. Prices are provider-specific and should be checked against each cited observation date.
Gemini 3.1 Flash-Lite's recorded capabilities are chat, generation, reasoning, and tools. Its supported tasks are audio, document, image, text, and video.
Gemini 3.1 Flash-Lite is not marked as an open-weight model. Self-hosting is marked as unsupported, with no license recorded.
No predecessor is recorded. No successor is recorded. Recorded aliases are gemini-3.1-flash-lite.
The current compatible peer set includes Gemini 2.5 Flash-Lite, Gemini 3 Flash, Nova 2 Lite, Qwen3.8-Flash, and Gemini Computer Use. Compatibility requires shared sourced modalities and tasks; it does not imply that one model is better. Browse model comparisons →
Gemini 3.1 Flash-Lite was last verified Aug 29, 2026 from Google DeepMind's primary source, “Gemini 3.1 Flash-Lite | Gemini API.” The evidence is classified as official fact.