AI Model Profile

Gemini 3.1 Flash-Lite

Google’s stable Gemini 3.1 Flash-Lite accepts text, images, video, audio and PDFs and produces text. This is a documentation review, not a Kingy performance test.

Family
Gemini 3
Release date
Unknown
Status
Stable
Context window
Input token limit: 1,048,576 (review overdue)
Output limit
65,536 tokens (review overdue)
API
yes
Open weights
Unknown
Local/self-hosted
Unknown
Pricing
Standard Gemini Developer API paid tier, USD per million tokens: text/image/video input $0.25; audio input $0.50; output including thinking $1.50. Other consumption modes and tool charges differ. Checked September 14, 2026.
Evidence state
Source-verified

Benchmark Caveat

Provider documentation and provider-published capability notes are directional, not universal rankings. Real results depend on prompts, tools, latency targets, pricing tier, safety filters, context length, and workload mix.

Best for

Evaluate for bounded extraction, translation and classification tasks; these are vendor-documented use cases, not measured Kingy recommendations.

Skip if

Skip the retired -preview identifier. Do not choose this model for native image/audio generation, computer use or the Live API; Google lists those capabilities as unsupported.

Strengths

Documented support for function calling, structured outputs and thinking. Input token limit: 1,048,576; output token limit: 65,536.

Weaknesses

No Kingy workload benchmark is published here. Account quotas, region eligibility, latency and task accuracy still need checks for your deployment.

Agent suitability

Function calling and structured output are documented. Validate tool permissions, schema conformance and failure handling in your own integration.

Kingy AI take

The stable model remains documented as available. Compare it with Google’s listed replacement, Gemini 3.5 Flash-Lite, on your workload; the retired preview is a different identifier.

Kingy AI Product Facts

Gemini 3.1 Flash-Lite

Current statusGenerally available

Company
Google
Primary job
Not yet reviewed
Audience
Not yet reviewed
API
Available
Evidence
Source-backed; review due
Coverage
0 of 88 core fields up to date
Sources
3
Latest source check
September 21, 2026
See all tracked product factsPricing, platforms, dependencies, data claims, regions, and timeline

Plans and pricing units

  • Text, image and video input 0.25 USD per 1,000,000 input tokens
  • Audio input 0.5 USD per 1,000,000 input tokens
  • Output, including thinking 1.5 USD per 1,000,000 output tokens

Where it runs

Availability
Not yet reviewed

Developer access

API access
Available

Stack and integrations

Model / provider dependencies
Not yet reviewed
Integrations
Not yet reviewed

Data and regions

Vendor data-use claim
Not yet reviewed
Vendor retention claim
Not yet reviewed
Regions
Not yet reviewed

Launch and latest material update

Launch
Not yet reviewed
Latest material update
Not yet reviewed
Technical evidence and revision history
Embed “Facts tracked by Kingy”

This label is not a security certification, audit, or product endorsement. It reports the facts Kingy currently supports and when they were last checked.

Full Model Notes

This record covers the stable API identifier gemini-3.1-flash-lite. It is distinct from gemini-3.1-flash-lite-preview, which Google marks as shut down.

Documentation reviewed September 14, 2026. Google’s deprecation schedule gives May 7, 2027 as the stable model’s earliest possible retirement date and lists Gemini 3.5 Flash-Lite as its replacement. That date is not a guaranteed exact shutdown time.

Documented API scope

The model documentation lists text, image, video, audio and PDF input; text output; an input limit of 1,048,576 tokens and an output limit of 65,536 tokens. Function calling, structured output and thinking are supported. Native image/audio generation, computer use and the Live API are not.

Price scope

Google’s Standard paid Gemini Developer API rates, in USD per million tokens, are $0.25 for text/image/video input, $0.50 for audio input and $1.50 for output including thinking. These rates exclude caching, grounding and other tool charges. Batch, Flex, Priority and free-tier rules differ; account and region eligibility must be checked separately.

Correction history

The September 14, 2026 correction separates the stable model from its retired preview across the profile and canonical fact records. It removes the prior instruction to avoid the stable model solely because the preview shut down. No API generation, latency or quality test was run for this correction. The original publication date and URL are preserved.

Coding notes

No coding-quality ranking established by this documentation review.

Reasoning notes

Thinking is documented; reasoning quality was not measured in this review.

Creative notes

Accepts multimodal inputs and outputs text. Native image and audio generation are not supported for this identifier.

Research notes

Search grounding and URL context are documented. Check citations and any separate tool charges.

API pricing notes

Standard paid API rates: $0.25 per 1M text/image/video input tokens, $0.50 per 1M audio input tokens, $1.50 per 1M output tokens including thinking. USD; excludes caching, grounding and other charges. See Google’s pricing source for Batch, Flex, Priority and free-tier scope.

License notes

This review establishes the documented hosted API route. It does not establish downloadable weights or redistribution rights; those fields remain unknown. Check the provider terms for your intended use.