AI Model Profile

Gemini 3.6 Flash

Gemini 3.6 Flash is Google's latest Flash-tier model, balancing speed with intelligence for agentic and multimodal tasks.

Family
Gemini 3
Release date
Unknown
Status
Stable
Context window
1,048,576 tokens
Output limit
65,536 tokens
API
yes
Open weights
no
Local/self-hosted
no
Pricing
Gemini 3.6 Flash paid-tier pricing checked 2026-07-28: $1.50 per million input tokens and $7.50 per million output tokens for the standard text workload represented in LANTERN-7.
Verification
verified

Verification & Sources

Status
Verified
Source links
5
Freshness
Verified July 28, 2026
Last verified
July 28, 2026
Last updated
July 28, 2026
Suggest a correction

Form submissions, correction notes, score details, URLs, and analytics events may be stored for editorial review, spam prevention, product improvement, and follow-up. Do not submit secrets, unreleased financials, private customer data, or regulated personal data through these forms.

Benchmark Caveat

Benchmarks and provider capability notes are directional, not universal rankings. Results can shift with prompts, tool use, latency targets, pricing tier, safety filters, context length, and the workload mix a real team runs.

See the linked official model, docs, model-card, or pricing source for provider-published capability notes.

Best for

Teams that need a very large context window across many input modalities at Flash-tier speed.

Skip if

Skip if you need published knowledge-cutoff dates for compliance, or if a Pro-tier model is required.

Strengths

Google lists Gemini 3.6 Flash with a 1,048,576-token input limit, a 65,536-token output limit, supported function calling, and the broadest input modality set in this directory: text, image, video, audio, and PDF. Generally available.

Weaknesses

Google does not publish a knowledge cutoff for this model. Output is text only, and the rate card varies by tier and modality.

Agent suitability

Useful for agent workflows when the provider supports tool use, long context, structured outputs, or workflow-specific APIs.

Kingy AI take

Use this as a source-backed shortlist candidate, not a universal ranking. Re-check official provider docs and run a task-specific trial before production adoption.

Coding notes

Use official docs and live evals before selecting this model for production coding workflows.

Reasoning notes

Provider capability notes are useful but should be validated on representative prompts and tools.

Creative notes

Use a small creative test set before standardizing outputs for brand, media, or customer-facing work.

Research notes

Track release notes and model lifecycle notices because availability and aliases can change.

API pricing notes

Use the stable GA ID gemini-3.6-flash with medium thinking and tools disabled. Record response modelVersion and usage metadata. Only the paid tier is eligible: Google's terms state paid content is not used to improve products; verify project-level zero-data-retention eligibility separately.

License notes

Commercial/API terms apply unless the linked official source states otherwise.

Hardware requirements

Cloud/API model; local hardware requirements are not published as a self-hosted path.