Kingy AI Launch Intelligence

AI Model Intelligence Hub

Kingy AI tracks AI models by provider, model family, modality, release status, API access, web access, local availability, open-weight/open-source status, pricing notes, hardware requirements, official sources, benchmark caveats, and last-verified status.

What Kingy AI Tracks

  • Provider, model family, modality, release status, access paths, pricing notes, hardware requirements, and last-verified status.
  • Open-weight/open-source status, license notes, official docs, model cards, source links, and benchmark caveats.
  • Use-case notes for coding, agents, local/private workflows, multimodal work, creators, and business teams.

Browse by Model Type

  • Explore open-weight, coding, local/private, multimodal, agent-ready, creator, and business-workflow model records.
  • Browse source-backed provider and family coverage for GPT, Claude, Gemini, Llama, Mistral, Qwen, DeepSeek, and other model lines.
Benchmark caveatBenchmarks are directional signals, not universal rankings. Results can shift with prompts, tool use, latency targets, pricing tier, eval contamination, safety filters, context length, and the task mix a real team runs.

Showing 1–22 of 22 models

Source-checked profiles within the review window appear first; retired models appear last. Check the date and limitations on each record. Indexing is assessed separately from review age.

Text

gpt-oss-20b

gpt-oss-20b is OpenAI's medium-sized open-weight gpt-oss model for low-latency, local, or specialized use cases.

API: No Open weights: Yes Local: Yes
Provider
OpenAI
Context
Unknown
Last verified
Needs recheck: checked June 24, 2026
Text

gpt-oss-120b

gpt-oss-120b is OpenAI's largest open-weight gpt-oss reasoning model, described by OpenAI as fitting into a single H100 GPU.

API: No Open weights: Yes Local: Yes
Provider
OpenAI
Context
Unknown
Last verified
Needs recheck: checked June 24, 2026
Text

Llama-3.1-Nemotron-Nano-8B-v1

Llama-3.1-Nemotron-Nano-8B-v1 is an NVIDIA compact open model that fits on a single RTX GPU and supports 128K context.

API: Yes Open weights: Yes Local: Yes
Provider
NVIDIA
Context
128K tokens
Last verified
Needs recheck: checked June 24, 2026
Text

Llama-3.3-Nemotron-Super-49B-v1.5

Llama-3.3-Nemotron-Super-49B-v1.5 is an NVIDIA reasoning model derived from Meta Llama 3.3 and tuned for RAG and tool calling.

API: Yes Open weights: Yes Local: Yes
Provider
NVIDIA
Context
128K tokens
Last verified
Needs recheck: checked June 24, 2026
Text

Ministral 3 3B

Ministral 3 3B is the smallest listed Ministral 3 variant for compact Mistral deployments.

API: Yes Open weights: Yes Local: Yes
Provider
Mistral AI
Context
256k tokens
Last verified
Needs recheck: checked June 24, 2026
Text

Ministral 3 8B

Ministral 3 8B is a compact Mistral model variant in the Ministral 3 family.

API: Yes Open weights: Yes Local: Yes
Provider
Mistral AI
Context
256k tokens
Last verified
Needs recheck: checked June 24, 2026
Text

Ministral 3 14B

Ministral 3 14B is a Mistral small-model variant listed with the Ministral 3 family in official documentation.

API: Yes Open weights: Yes Local: Yes
Provider
Mistral AI
Context
256k tokens
Last verified
Needs recheck: checked June 24, 2026
Text

NVIDIA Nemotron 3 Super 120B-A12B

NVIDIA Nemotron 3 Super 120B-A12B is part of NVIDIA's Nemotron open-model family for agentic AI and reasoning workflows.

API: Yes Open weights: Yes Local: Yes
Provider
NVIDIA
Context
Unknown
Last verified
Needs recheck: checked June 24, 2026
Text

Phi-4

Phi-4 is a Microsoft Research open model carded on Hugging Face for high-quality reasoning-focused tasks.

API: Yes Open weights: Yes Local: Yes
Provider
Microsoft
Context
Unknown
Last verified
Needs recheck: checked June 24, 2026
Text

Phi-4-mini-instruct

Phi-4-mini-instruct is a lightweight Microsoft open model in the Phi-4 family with 128K context listed on its model card.

API: Yes Open weights: Yes Local: Yes
Provider
Microsoft
Context
128K tokens
Last verified
Needs recheck: checked June 24, 2026
Text

Phi-4-mini-reasoning

Phi-4-mini-reasoning is a lightweight Microsoft open model for advanced math reasoning in the Phi-4 family.

API: Yes Open weights: Yes Local: Yes
Provider
Microsoft
Context
128K tokens
Last verified
Needs recheck: checked June 24, 2026
Audio, Image, Multimodal, Text

Phi-4-multimodal-instruct

Phi-4-multimodal-instruct is a Microsoft lightweight multimodal foundation model for text, image, and audio inputs with text output.

API: Yes Open weights: Yes Local: Yes
Provider
Microsoft
Context
128K tokens
Last verified
Needs recheck: checked June 24, 2026
Text

Phi-4-reasoning

Phi-4-reasoning is a Microsoft open-weight reasoning model finetuned from Phi-4 for math, science, and coding skills.

API: Yes Open weights: Yes Local: Yes
Provider
Microsoft
Context
Unknown
Last verified
Needs recheck: checked June 24, 2026
Text

Qwen3-4B

Qwen3-4B is a small dense Qwen3 open-weight model listed by Qwen.

API: Yes Open weights: Yes Local: Yes
Provider
Qwen
Context
32K tokens
Last verified
Needs recheck: checked June 24, 2026
Text

Qwen3-8B

Qwen3-8B is a compact dense Qwen3 open-weight model listed in Qwen's official release.

API: Yes Open weights: Yes Local: Yes
Provider
Qwen
Context
32K tokens
Last verified
Needs recheck: checked June 24, 2026
Text

Qwen3-14B

Qwen3-14B is a dense open-weight Qwen3 model listed by Qwen.

API: Yes Open weights: Yes Local: Yes
Provider
Qwen
Context
32K tokens
Last verified
Needs recheck: checked June 24, 2026
Text

Qwen3-30B-A3B

Qwen3-30B-A3B is a Qwen3 mixture-of-experts open-weight model described in Qwen's official Qwen3 release.

API: Yes Open weights: Yes Local: Yes
Provider
Qwen
Context
32K tokens
Last verified
Needs recheck: checked June 24, 2026
Text

Qwen3-32B

Qwen3-32B is a dense open-weight Qwen3 model listed in the official Qwen3 release.

API: Yes Open weights: Yes Local: Yes
Provider
Qwen
Context
32K tokens
Last verified
Needs recheck: checked June 24, 2026
Image

Stable Diffusion 3.5 Large

Stable Diffusion 3.5 Large is Stability AI's open text-to-image model for professional image generation and customization.

API: Yes Open weights: Yes Local: Yes
Provider
Stability AI
Context
Unknown
Last verified
Needs recheck: checked June 24, 2026
Image

Stable Diffusion 3.5 Large Turbo

Stable Diffusion 3.5 Large Turbo is Stability AI's distilled SD3.5 Large variant for faster image generation with fewer steps.

API: Yes Open weights: Yes Local: Yes
Provider
Stability AI
Context
Unknown
Last verified
Needs recheck: checked June 24, 2026
Image, Multimodal, Text

Llama 4 Scout

Llama 4 Scout is a Meta Llama 4 mixture-of-experts model optimized for multimodal understanding, multilingual tasks, coding, tool-calling, and agentic systems.

API: Unknown Open weights: Yes Local: Yes
Provider
Meta
Context
10M tokens
Last verified
Needs recheck: checked June 18, 2026
Text

Qwen3-235B-A22B

Qwen3-235B-A22B is a flagship Qwen3 mixture-of-experts model described in Qwen release materials as open-weight and useful across coding, math, general capabilities, and agent workflows.

API: Unknown Open weights: Yes Local: Yes
Provider
Qwen
Context
128K tokens
Last verified
Needs recheck: checked June 18, 2026