Kingy AI Launch Intelligence

AI Model Intelligence Hub

Kingy AI tracks AI models by provider, model family, modality, release status, API access, web access, local availability, open-weight/open-source status, pricing notes, hardware requirements, official sources, benchmark caveats, and last-verified status.

What Kingy AI Tracks

  • Provider, model family, modality, release status, access paths, pricing notes, hardware requirements, and last-verified status.
  • Open-weight/open-source status, license notes, official docs, model cards, source links, and benchmark caveats.
  • Use-case notes for coding, agents, local/private workflows, multimodal work, creators, and business teams.

Browse by Model Type

  • Explore open-weight, coding, local/private, multimodal, agent-ready, creator, and business-workflow model records.
  • Browse source-backed provider and family coverage for GPT, Claude, Gemini, Llama, Mistral, Qwen, DeepSeek, and other model lines.
Benchmark caveatBenchmarks are directional signals, not universal rankings. Results can shift with prompts, tool use, latency targets, pricing tier, eval contamination, safety filters, context length, and the task mix a real team runs.

Showing 1–24 of 45 models

Source-checked profiles within the review window appear first; retired models appear last. Check the date and limitations on each record. Indexing is assessed separately from review age.

Image, Multimodal, Text

Gemini 2.5 Pro

Gemini 2.5 Pro is a Google Gemini API model described as advanced for complex tasks, deep reasoning, and coding capabilities.

API: Yes Open weights: No Local: No
Provider
Google
Context
Unknown
Last verified
Needs recheck: checked August 16, 2026
Multimodal, Text

Claude Fable 5

Claude Fable 5 is Anthropic's most capable widely released model, listed for demanding reasoning and long-horizon agentic work.

API: Yes Open weights: No Local: No
Provider
Anthropic
Context
1M tokens
Last verified
Needs recheck: checked July 28, 2026
Text

DeepSeek V4 Pro

DeepSeek V4 Pro is listed in DeepSeek API docs as a current model supporting thinking and non-thinking modes, JSON output, tool calls, and a 1M context length.

API: Yes Open weights: No Local: No
Provider
DeepSeek
Context
1M tokens
Last verified
Needs recheck: checked July 28, 2026
Image, Multimodal, Text

Gemini 3.6 Flash

Gemini 3.6 Flash is Google's latest Flash-tier model, balancing speed with intelligence for agentic and multimodal tasks.

API: Yes Open weights: No Local: No
Provider
Google
Context
1,048,576 tokens
Last verified
Needs recheck: checked July 28, 2026
Image, Multimodal, Text

Mistral Medium 3.5

Mistral Medium 3.5 is listed by Mistral as a frontier-class multimodal model optimized for agentic and coding use cases.

API: Yes Open weights: Yes Local: Yes
Provider
Mistral AI
Context
256K tokens
Last verified
Needs recheck: checked July 28, 2026
Multimodal, Text

Claude Opus 5

Claude Opus 5 is Anthropic's recommended default model for complex agentic coding and enterprise work, with a 1M-token context window and 128K maximum output.

API: Yes Open weights: No Local: No
Provider
Anthropic
Context
1M tokens
Last verified
Needs recheck: checked July 25, 2026
Multimodal, Text

Claude Sonnet 5

Claude Sonnet 5 is Anthropic's speed-and-intelligence balance in the Claude 5 line, with a 1M-token context window and 128K maximum output.

API: Yes Open weights: No Local: No
Provider
Anthropic
Context
1M tokens
Last verified
Needs recheck: checked July 25, 2026
Image, Text

GPT-5.6 Luna

GPT-5.6 Luna is the cost-optimised tier of the GPT-5.6 family, retaining the full 1,050,000-token context window and reasoning-token support.

API: Yes Open weights: No Local: No
Provider
OpenAI
Context
1,050,000 tokens
Last verified
Needs recheck: checked July 25, 2026
Image, Text

GPT-5.6 Sol

GPT-5.6 Sol is OpenAI's frontier model for complex professional work, and the model the bare gpt-5.6 alias resolves to.

API: Yes Open weights: No Local: No
Provider
OpenAI
Context
1,050,000 tokens
Last verified
Needs recheck: checked July 25, 2026
Image, Text

GPT-5.6 Terra

GPT-5.6 Terra is the balanced intelligence-and-cost tier of the GPT-5.6 family, positioned roughly where the mini tier sat in earlier GPT-5 generations.

API: Yes Open weights: No Local: No
Provider
OpenAI
Context
1,050,000 tokens
Last verified
Needs recheck: checked July 25, 2026
Code, Text

Codestral

Codestral is Mistral AI's coding-focused model line, listed in Mistral model documentation and context-window notes.

API: Yes Open weights: No Local: No
Provider
Mistral AI
Context
128k tokens
Last verified
Needs recheck: checked June 24, 2026
Multimodal, Text

GPT-5.4 Pro

GPT-5.4 Pro is an OpenAI model catalog entry for teams that want a smarter GPT-5.4 variant for professional workflows.

API: Yes Open weights: No Local: No
Provider
OpenAI
Context
1M tokens
Last verified
Needs recheck: checked June 24, 2026
Multimodal, Text

GPT-5.5 Pro

GPT-5.5 Pro is an OpenAI frontier model variant listed in the OpenAI model catalog for higher-precision professional and coding work.

API: Yes Open weights: No Local: No
Provider
OpenAI
Context
1M tokens
Last verified
Needs recheck: checked June 24, 2026
Text

gpt-oss-120b

gpt-oss-120b is OpenAI's largest open-weight gpt-oss reasoning model, described by OpenAI as fitting into a single H100 GPU.

API: No Open weights: Yes Local: Yes
Provider
OpenAI
Context
Unknown
Last verified
Needs recheck: checked June 24, 2026
Text

Phi-4

Phi-4 is a Microsoft Research open model carded on Hugging Face for high-quality reasoning-focused tasks.

API: Yes Open weights: Yes Local: Yes
Provider
Microsoft
Context
Unknown
Last verified
Needs recheck: checked June 24, 2026
Text

Phi-4-mini-instruct

Phi-4-mini-instruct is a lightweight Microsoft open model in the Phi-4 family with 128K context listed on its model card.

API: Yes Open weights: Yes Local: Yes
Provider
Microsoft
Context
128K tokens
Last verified
Needs recheck: checked June 24, 2026
Text

Phi-4-reasoning

Phi-4-reasoning is a Microsoft open-weight reasoning model finetuned from Phi-4 for math, science, and coding skills.

API: Yes Open weights: Yes Local: Yes
Provider
Microsoft
Context
Unknown
Last verified
Needs recheck: checked June 24, 2026
Text

Qwen3-8B

Qwen3-8B is a compact dense Qwen3 open-weight model listed in Qwen's official release.

API: Yes Open weights: Yes Local: Yes
Provider
Qwen
Context
32K tokens
Last verified
Needs recheck: checked June 24, 2026
Text

Qwen3-14B

Qwen3-14B is a dense open-weight Qwen3 model listed by Qwen.

API: Yes Open weights: Yes Local: Yes
Provider
Qwen
Context
32K tokens
Last verified
Needs recheck: checked June 24, 2026
Text

Qwen3-30B-A3B

Qwen3-30B-A3B is a Qwen3 mixture-of-experts open-weight model described in Qwen's official Qwen3 release.

API: Yes Open weights: Yes Local: Yes
Provider
Qwen
Context
32K tokens
Last verified
Needs recheck: checked June 24, 2026
Text

Qwen3-32B

Qwen3-32B is a dense open-weight Qwen3 model listed in the official Qwen3 release.

API: Yes Open weights: Yes Local: Yes
Provider
Qwen
Context
32K tokens
Last verified
Needs recheck: checked June 24, 2026
Text

DeepSeek V4 Flash

DeepSeek V4 Flash is listed in DeepSeek API docs as a current model supporting thinking and non-thinking modes, JSON output, tool calls, and a 1M context length.

API: Yes Open weights: Yes Local: Yes
Provider
DeepSeek
Context
1,000,000 tokens (review overdue)
Last verified
Needs recheck: checked June 18, 2026
Text

Devstral 2

Devstral 2 is listed by Mistral as a frontier code agents model for software engineering tasks.

API: Yes Open weights: Unknown Local: Unknown
Provider
Mistral AI
Context
256K tokens
Last verified
Needs recheck: checked June 18, 2026
Image, Multimodal, Text

Gemini 3 Flash

Gemini 3 Flash is a Google Gemini API preview model positioned as frontier-class performance at a lower cost tier than larger models.

API: Yes Open weights: No Local: No
Provider
Google
Context
Unknown
Last verified
Needs recheck: checked June 18, 2026