Kingy AI Launch Intelligence

AI Model Intelligence Hub

Kingy AI tracks AI models by provider, model family, modality, release status, API access, web access, local availability, open-weight/open-source status, pricing notes, hardware requirements, official sources, benchmark caveats, and last-verified status.

What Kingy AI Tracks

  • Provider, model family, modality, release status, access paths, pricing notes, hardware requirements, and last-verified status.
  • Open-weight/open-source status, license notes, official docs, model cards, source links, and benchmark caveats.
  • Use-case notes for coding, agents, local/private workflows, multimodal work, creators, and business teams.

Browse by Model Type

  • Explore open-weight, coding, local/private, multimodal, agent-ready, creator, and business-workflow model records.
  • Browse source-backed provider and family coverage for GPT, Claude, Gemini, Llama, Mistral, Qwen, DeepSeek, and other model lines.
Benchmark caveatBenchmarks are directional signals, not universal rankings. Results can shift with prompts, tool use, latency targets, pricing tier, eval contamination, safety filters, context length, and the task mix a real team runs.

Showing 25–48 of 51 models

Source-checked profiles within the review window appear first; retired models appear last. Check the date and limitations on each record. Indexing is assessed separately from review age.

Text

Llama-3.1-Nemotron-Nano-8B-v1

Llama-3.1-Nemotron-Nano-8B-v1 is an NVIDIA compact open model that fits on a single RTX GPU and supports 128K context.

API: Yes Open weights: Yes Local: Yes
Provider
NVIDIA
Context
128K tokens
Last verified
Needs recheck: checked June 24, 2026
Text

Llama-3.3-Nemotron-Super-49B-v1.5

Llama-3.3-Nemotron-Super-49B-v1.5 is an NVIDIA reasoning model derived from Meta Llama 3.3 and tuned for RAG and tool calling.

API: Yes Open weights: Yes Local: Yes
Provider
NVIDIA
Context
128K tokens
Last verified
Needs recheck: checked June 24, 2026
Text

Ministral 3 14B

Ministral 3 14B is a Mistral small-model variant listed with the Ministral 3 family in official documentation.

API: Yes Open weights: Yes Local: Yes
Provider
Mistral AI
Context
256k tokens
Last verified
Needs recheck: checked June 24, 2026
Text

NVIDIA Nemotron 3 Super 120B-A12B

NVIDIA Nemotron 3 Super 120B-A12B is part of NVIDIA's Nemotron open-model family for agentic AI and reasoning workflows.

API: Yes Open weights: Yes Local: Yes
Provider
NVIDIA
Context
Unknown
Last verified
Needs recheck: checked June 24, 2026
Text

Qwen3-30B-A3B

Qwen3-30B-A3B is a Qwen3 mixture-of-experts open-weight model described in Qwen's official Qwen3 release.

API: Yes Open weights: Yes Local: Yes
Provider
Qwen
Context
32K tokens
Last verified
Needs recheck: checked June 24, 2026
Image, Text

Claude Haiku 4.5

Claude Haiku 4.5 is Anthropic's Haiku-tier model described in official docs as the fastest current model with near-frontier intelligence.

API: Yes Open weights: No Local: No
Provider
Anthropic
Context
200K tokens
Last verified
Needs recheck: checked June 18, 2026
Text

DeepSeek V4 Flash

DeepSeek V4 Flash is listed in DeepSeek API docs as a current model supporting thinking and non-thinking modes, JSON output, tool calls, and a 1M context length.

API: Yes Open weights: Yes Local: Yes
Provider
DeepSeek
Context
1,000,000 tokens (review overdue)
Last verified
Needs recheck: checked June 18, 2026
Text

Devstral 2

Devstral 2 is listed by Mistral as a frontier code agents model for software engineering tasks.

API: Yes Open weights: Unknown Local: Unknown
Provider
Mistral AI
Context
256K tokens
Last verified
Needs recheck: checked June 18, 2026
Image, Multimodal, Text

Gemini 3 Flash

Gemini 3 Flash is a Google Gemini API preview model positioned as frontier-class performance at a lower cost tier than larger models.

API: Yes Open weights: No Local: No
Provider
Google
Context
Unknown
Last verified
Needs recheck: checked June 18, 2026
Image, Multimodal, Text

Gemini 3.1 Pro

Gemini 3.1 Pro is a Google Gemini API preview model described for advanced intelligence, complex problem-solving, and agentic or vibe-coding capabilities.

API: Yes Open weights: No Local: No
Provider
Google
Context
Unknown
Last verified
Needs recheck: checked June 18, 2026
Image, Multimodal, Text

Gemini 3.5 Flash

Gemini 3.5 Flash is a Google Gemini API model listed as stable and positioned for sustained frontier performance on agentic and coding tasks.

API: Yes Open weights: No Local: No
Provider
Google
Context
Unknown
Last verified
Needs recheck: checked June 18, 2026
Image, Text

GPT-5.4

GPT-5.4 is an OpenAI API model described in official docs as a more affordable model for coding and professional work.

API: Yes Open weights: No Local: No
Provider
OpenAI
Context
1M tokens
Last verified
Needs recheck: checked June 18, 2026
Image, Text

GPT-5.4 mini

GPT-5.4 mini is an OpenAI API model described as a smaller model for coding, computer use, and subagents.

API: Yes Open weights: No Local: No
Provider
OpenAI
Context
400K tokens
Last verified
Needs recheck: checked June 18, 2026
Image, Text

GPT-5.5

GPT-5.5 is OpenAI's flagship API model for complex reasoning, coding, and professional work, with text and image input and text output described in official OpenAI model docs.

API: Yes Open weights: No Local: No
Provider
OpenAI
Context
1M tokens
Last verified
Needs recheck: checked June 18, 2026
Image, Text

Grok 4.3

Grok 4.3 is an xAI API model listed with agentic tool calling, configurable reasoning, and a 1M-token context window.

API: Yes Open weights: No Local: No
Provider
xAI
Context
1M tokens
Last verified
Needs recheck: checked June 18, 2026
Text

Grok Build 0.1

Grok Build 0.1 is an xAI coding model described as trained specifically for fast agentic coding workflows.

API: Yes Open weights: No Local: No
Provider
xAI
Context
256K tokens
Last verified
Needs recheck: checked June 18, 2026
Image, Multimodal, Text

Llama 4 Maverick

Llama 4 Maverick is a larger Meta Llama 4 mixture-of-experts model optimized for multimodal understanding, coding, tool-calling, and agentic systems.

API: Unknown Open weights: Yes Local: Yes
Provider
Meta
Context
1M tokens
Last verified
Needs recheck: checked June 18, 2026
Image, Multimodal, Text

Llama 4 Scout

Llama 4 Scout is a Meta Llama 4 mixture-of-experts model optimized for multimodal understanding, multilingual tasks, coding, tool-calling, and agentic systems.

API: Unknown Open weights: Yes Local: Yes
Provider
Meta
Context
10M tokens
Last verified
Needs recheck: checked June 18, 2026
Text

Qwen3-235B-A22B

Qwen3-235B-A22B is a flagship Qwen3 mixture-of-experts model described in Qwen release materials as open-weight and useful across coding, math, general capabilities, and agent workflows.

API: Unknown Open weights: Yes Local: Yes
Provider
Qwen
Context
128K tokens
Last verified
Needs recheck: checked June 18, 2026
Multimodal, Text

Claude Opus 4.6

Claude Opus 4.6 is an Anthropic Opus-tier model included in Claude pricing, data residency, and long-context notes.

API: Yes Open weights: No Local: No
Provider
Anthropic
Context
1M tokens
Last verified
Needs recheck: checked June 24, 2026
Multimodal, Text

Claude Opus 4.7

Claude Opus 4.7 is an Anthropic Opus-tier model listed in official Claude pricing and lifecycle documentation.

API: Yes Open weights: No Local: No
Provider
Anthropic
Context
1M tokens
Last verified
Needs recheck: checked June 24, 2026
Multimodal, Text

Claude Sonnet 4.5

Claude Sonnet 4.5 is a Sonnet-tier Claude model listed by Anthropic with current pricing and cloud-platform notes.

API: Yes Open weights: No Local: No
Provider
Anthropic
Context
200K tokens
Last verified
Needs recheck: checked June 24, 2026
Text

DeepSeek-V3

DeepSeek-V3 is a DeepSeek model release described by DeepSeek as improving speed and API compatibility.

API: Yes Open weights: Yes Local: Yes
Provider
DeepSeek
Context
Unknown
Last verified
Needs recheck: checked June 24, 2026
Text

DeepSeek-V3.2

DeepSeek-V3.2 is a DeepSeek model release described by DeepSeek as integrating thinking into tool-use.

API: Yes Open weights: Yes Local: Yes
Provider
DeepSeek
Context
Unknown
Last verified
Needs recheck: checked June 24, 2026