Kingy AI Launch Intelligence

AI Model Intelligence Hub

Kingy AI tracks AI models by provider, model family, modality, release status, API access, web access, local availability, open-weight/open-source status, pricing notes, hardware requirements, official sources, benchmark caveats, and last-verified status.

What Kingy AI Tracks

  • Provider, model family, modality, release status, access paths, pricing notes, hardware requirements, and last-verified status.
  • Open-weight/open-source status, license notes, official docs, model cards, source links, and benchmark caveats.
  • Use-case notes for coding, agents, local/private workflows, multimodal work, creators, and business teams.

Browse by Model Type

  • Explore open-weight, coding, local/private, multimodal, agent-ready, creator, and business-workflow model records.
  • Browse source-backed provider and family coverage for GPT, Claude, Gemini, Llama, Mistral, Qwen, DeepSeek, and other model lines.
Benchmark caveatBenchmarks are directional signals, not universal rankings. Results can shift with prompts, tool use, latency targets, pricing tier, eval contamination, safety filters, context length, and the task mix a real team runs.

Showing 1–24 of 51 models

Source-checked profiles within the review window appear first; retired models appear last. Check the date and limitations on each record. Indexing is assessed separately from review age.

Image, Multimodal, Text

Gemini 3.1 Flash-Lite

Google’s stable Gemini 3.1 Flash-Lite accepts text, images, video, audio and PDFs and produces text. This is a documentation review, not a Kingy performance test.

API: Yes Open weights: Unknown Local: Unknown
Provider
Google
Context
Input token limit: 1,048,576 (review overdue)
Last verified
2026-09-14
Image, Multimodal, Text

Gemini 2.5 Flash

Gemini 2.5 Flash is a Google Gemini API model described as price-performance oriented for low-latency, high-volume tasks that require reasoning.

API: Yes Open weights: No Local: No
Provider
Google
Context
Unknown
Last verified
Needs recheck: checked August 16, 2026
Multimodal, Text

Gemini 2.5 Flash-Lite

Gemini 2.5 Flash-Lite is listed by Google as the fastest and most budget-friendly multimodal model in the Gemini 2.5 family.

API: Yes Open weights: No Local: No
Provider
Google
Context
Unknown
Last verified
Needs recheck: checked August 16, 2026
Multimodal, Text

Claude Fable 5

Claude Fable 5 is Anthropic's most capable widely released model, listed for demanding reasoning and long-horizon agentic work.

API: Yes Open weights: No Local: No
Provider
Anthropic
Context
1M tokens
Last verified
Needs recheck: checked July 28, 2026
Text

DeepSeek V4 Pro

DeepSeek V4 Pro is listed in DeepSeek API docs as a current model supporting thinking and non-thinking modes, JSON output, tool calls, and a 1M context length.

API: Yes Open weights: No Local: No
Provider
DeepSeek
Context
1M tokens
Last verified
Needs recheck: checked July 28, 2026
Image, Multimodal, Text

Gemini 3.6 Flash

Gemini 3.6 Flash is Google's latest Flash-tier model, balancing speed with intelligence for agentic and multimodal tasks.

API: Yes Open weights: No Local: No
Provider
Google
Context
1,048,576 tokens
Last verified
Needs recheck: checked July 28, 2026
Image, Multimodal, Text

Mistral Medium 3.5

Mistral Medium 3.5 is listed by Mistral as a frontier-class multimodal model optimized for agentic and coding use cases.

API: Yes Open weights: Yes Local: Yes
Provider
Mistral AI
Context
256K tokens
Last verified
Needs recheck: checked July 28, 2026
Multimodal, Text

Claude Opus 5

Claude Opus 5 is Anthropic's recommended default model for complex agentic coding and enterprise work, with a 1M-token context window and 128K maximum output.

API: Yes Open weights: No Local: No
Provider
Anthropic
Context
1M tokens
Last verified
Needs recheck: checked July 25, 2026
Multimodal, Text

Claude Sonnet 5

Claude Sonnet 5 is Anthropic's speed-and-intelligence balance in the Claude 5 line, with a 1M-token context window and 128K maximum output.

API: Yes Open weights: No Local: No
Provider
Anthropic
Context
1M tokens
Last verified
Needs recheck: checked July 25, 2026
Image, Text

GPT-5.6 Luna

GPT-5.6 Luna is the cost-optimised tier of the GPT-5.6 family, retaining the full 1,050,000-token context window and reasoning-token support.

API: Yes Open weights: No Local: No
Provider
OpenAI
Context
1,050,000 tokens
Last verified
Needs recheck: checked July 25, 2026
Image, Text

GPT-5.6 Sol

GPT-5.6 Sol is OpenAI's frontier model for complex professional work, and the model the bare gpt-5.6 alias resolves to.

API: Yes Open weights: No Local: No
Provider
OpenAI
Context
1,050,000 tokens
Last verified
Needs recheck: checked July 25, 2026
Image, Text

GPT-5.6 Terra

GPT-5.6 Terra is the balanced intelligence-and-cost tier of the GPT-5.6 family, positioned roughly where the mini tier sat in earlier GPT-5 generations.

API: Yes Open weights: No Local: No
Provider
OpenAI
Context
1,050,000 tokens
Last verified
Needs recheck: checked July 25, 2026
Multimodal, Text

Claude Mythos 5

Claude Mythos 5 is an Anthropic Project Glasswing model listed as limited availability for approved customers.

API: Yes Open weights: No Local: No
Provider
Anthropic
Context
1M tokens
Last verified
Needs recheck: checked June 24, 2026
Code, Text

Codestral

Codestral is Mistral AI's coding-focused model line, listed in Mistral model documentation and context-window notes.

API: Yes Open weights: No Local: No
Provider
Mistral AI
Context
128k tokens
Last verified
Needs recheck: checked June 24, 2026
Text

Command A

Command A is Cohere's enterprise model for tool use, RAG, agents, and multilingual work, listed with 111B parameters and 256K context.

API: Yes Open weights: No Local: No
Provider
Cohere
Context
256K tokens
Last verified
Needs recheck: checked June 24, 2026
Multimodal, Text

Command A+

Command A+ is Cohere's Command A family model announced with vision inputs, reasoning, translation, and agentic task support.

API: Yes Open weights: No Local: No
Provider
Cohere
Context
Unknown
Last verified
Needs recheck: checked June 24, 2026
Text

Command R7B

Command R7B is Cohere's smallest and fastest R-family model for RAG, tool use, and agents.

API: Yes Open weights: No Local: No
Provider
Cohere
Context
128K tokens
Last verified
Needs recheck: checked June 24, 2026
Text

Command R Plus

Command R+ is Cohere's R-family model optimized for conversational interaction, long-context tasks, complex RAG, and multi-step tool use.

API: Yes Open weights: No Local: No
Provider
Cohere
Context
128K tokens
Last verified
Needs recheck: checked June 24, 2026
Multimodal, Text

Gemini 2.5 Computer Use

Gemini 2.5 Computer Use is Google's agentic model for browser automation, UI testing, and visual interface reasoning.

API: Yes Open weights: No Local: No
Provider
Google
Context
Unknown
Last verified
Needs recheck: checked June 24, 2026
Multimodal, Text

GPT-5.4 nano

GPT-5.4 nano is listed by OpenAI as the cheapest GPT-5.4-class model for simple high-volume tasks.

API: Yes Open weights: No Local: No
Provider
OpenAI
Context
1M tokens
Last verified
Needs recheck: checked June 24, 2026
Multimodal, Text

GPT-5.4 Pro

GPT-5.4 Pro is an OpenAI model catalog entry for teams that want a smarter GPT-5.4 variant for professional workflows.

API: Yes Open weights: No Local: No
Provider
OpenAI
Context
1M tokens
Last verified
Needs recheck: checked June 24, 2026
Multimodal, Text

GPT-5.5 Pro

GPT-5.5 Pro is an OpenAI frontier model variant listed in the OpenAI model catalog for higher-precision professional and coding work.

API: Yes Open weights: No Local: No
Provider
OpenAI
Context
1M tokens
Last verified
Needs recheck: checked June 24, 2026
Text

gpt-oss-20b

gpt-oss-20b is OpenAI's medium-sized open-weight gpt-oss model for low-latency, local, or specialized use cases.

API: No Open weights: Yes Local: Yes
Provider
OpenAI
Context
Unknown
Last verified
Needs recheck: checked June 24, 2026
Text

gpt-oss-120b

gpt-oss-120b is OpenAI's largest open-weight gpt-oss reasoning model, described by OpenAI as fitting into a single H100 GPU.

API: No Open weights: Yes Local: Yes
Provider
OpenAI
Context
Unknown
Last verified
Needs recheck: checked June 24, 2026