AI Model Profile
Gemini 3.1 Flash-Lite
Google’s stable Gemini 3.1 Flash-Lite accepts text, images, video, audio and PDFs and produces text. This is a documentation review, not a Kingy performance test.
- Family
- Gemini 3
- Release date
- Unknown
- Status
- Stable
- Context window
- Input token limit: 1,048,576 (review overdue)
- Output limit
- 65,536 tokens (review overdue)
- API
- yes
- Open weights
- Unknown
- Local/self-hosted
- Unknown
- Pricing
- Standard Gemini Developer API paid tier, USD per million tokens: text/image/video input $0.25; audio input $0.50; output including thinking $1.50. Other consumption modes and tool charges differ. Checked September 14, 2026.
- Evidence state
- Source-verified
Benchmark Caveat
Provider documentation and provider-published capability notes are directional, not universal rankings. Real results depend on prompts, tools, latency targets, pricing tier, safety filters, context length, and workload mix.
Best for
Evaluate for bounded extraction, translation and classification tasks; these are vendor-documented use cases, not measured Kingy recommendations.
Skip if
Skip the retired -preview identifier. Do not choose this model for native image/audio generation, computer use or the Live API; Google lists those capabilities as unsupported.
Strengths
Documented support for function calling, structured outputs and thinking. Input token limit: 1,048,576; output token limit: 65,536.
Weaknesses
No Kingy workload benchmark is published here. Account quotas, region eligibility, latency and task accuracy still need checks for your deployment.
Agent suitability
Function calling and structured output are documented. Validate tool permissions, schema conformance and failure handling in your own integration.
Kingy AI take
The stable model remains documented as available. Compare it with Google’s listed replacement, Gemini 3.5 Flash-Lite, on your workload; the retired preview is a different identifier.
Kingy AI Product Facts
Gemini 3.1 Flash-Lite
Current statusGenerally available
- Company
- Primary job
- Not yet reviewed
- Audience
- Not yet reviewed
- API
- Available
- Evidence
- Source-backed; review due
- Coverage
- 0 of 88 core fields up to date
- Sources
- 3
- Latest source check
- September 21, 2026
See all tracked product factsPricing, platforms, dependencies, data claims, regions, and timeline
Plans and pricing units
- Text, image and video input 0.25 USD per 1,000,000 input tokens
- Audio input 0.5 USD per 1,000,000 input tokens
- Output, including thinking 1.5 USD per 1,000,000 output tokens
Where it runs
- Availability
- Not yet reviewed
Developer access
- API access
- Available
Stack and integrations
- Model / provider dependencies
- Not yet reviewed
- Integrations
- Not yet reviewed
Data and regions
- Vendor data-use claim
- Not yet reviewed
- Vendor retention claim
- Not yet reviewed
- Regions
- Not yet reviewed
Launch and latest material update
- Launch
- Not yet reviewed
- Latest material update
- Not yet reviewed
Full Model Notes
This record covers the stable API identifier gemini-3.1-flash-lite. It is distinct from gemini-3.1-flash-lite-preview, which Google marks as shut down.
Documentation reviewed September 14, 2026. Google’s deprecation schedule gives May 7, 2027 as the stable model’s earliest possible retirement date and lists Gemini 3.5 Flash-Lite as its replacement. That date is not a guaranteed exact shutdown time.
Documented API scope
The model documentation lists text, image, video, audio and PDF input; text output; an input limit of 1,048,576 tokens and an output limit of 65,536 tokens. Function calling, structured output and thinking are supported. Native image/audio generation, computer use and the Live API are not.
Price scope
Google’s Standard paid Gemini Developer API rates, in USD per million tokens, are $0.25 for text/image/video input, $0.50 for audio input and $1.50 for output including thinking. These rates exclude caching, grounding and other tool charges. Batch, Flex, Priority and free-tier rules differ; account and region eligibility must be checked separately.
Correction history
The September 14, 2026 correction separates the stable model from its retired preview across the profile and canonical fact records. It removes the prior instruction to avoid the stable model solely because the preview shut down. No API generation, latency or quality test was run for this correction. The original publication date and URL are preserved.
The Kingy Brief
Get the next Kingy Brief.
Source-checked AI changes, original tests and one practical thing to try.
Free · Choose your subjects · Double opt-in · Unsubscribe anytime
Coding notes
No coding-quality ranking established by this documentation review.
Reasoning notes
Thinking is documented; reasoning quality was not measured in this review.
Creative notes
Accepts multimodal inputs and outputs text. Native image and audio generation are not supported for this identifier.
Research notes
Search grounding and URL context are documented. Check citations and any separate tool charges.
API pricing notes
Standard paid API rates: $0.25 per 1M text/image/video input tokens, $0.50 per 1M audio input tokens, $1.50 per 1M output tokens including thinking. USD; excludes caching, grounding and other charges. See Google’s pricing source for Batch, Flex, Priority and free-tier scope.
License notes
This review establishes the documented hosted API route. It does not establish downloadable weights or redistribution rights; those fields remain unknown. Check the provider terms for your intended use.
Official Model Links
Model Intelligence Research Map
Use these internal paths to move from this model profile into provider pages, static comparison pages, related Kingy records, and the broader AI launch graph. These are research paths, not rankings.