Kingy AI Launch Intelligence

AI Model Intelligence Hub

Kingy AI tracks AI models by provider, model family, modality, release status, API access, web access, local availability, open-weight/open-source status, pricing notes, hardware requirements, official sources, benchmark caveats, and last-verified status.

What Kingy AI Tracks

  • Provider, model family, modality, release status, access paths, pricing notes, hardware requirements, and last-verified status.
  • Open-weight/open-source status, license notes, official docs, model cards, source links, and benchmark caveats.
  • Use-case notes for coding, agents, local/private workflows, multimodal work, creators, and business teams.

Browse by Model Type

  • Explore open-weight, coding, local/private, multimodal, agent-ready, creator, and business-workflow model records.
  • Browse source-backed provider and family coverage for GPT, Claude, Gemini, Llama, Mistral, Qwen, DeepSeek, and other model lines.
Benchmark caveatBenchmarks are directional signals, not universal rankings. Results can shift with prompts, tool use, latency targets, pricing tier, eval contamination, safety filters, context length, and the task mix a real team runs.

Showing 25–48 of 51 models

Fresh, verified profiles appear first. Records awaiting re-verification remain available in the research queue and are not indexed.

Text

Devstral 2

Devstral 2 is listed by Mistral as a frontier code agents model for software engineering tasks.

API: Yes Open weights: Unknown Local: Unknown
Provider
Mistral AI
Context
256K tokens
Last verified
Needs recheck: verified June 18, 2026
Multimodal, Text

Gemini 2.5 Computer Use

Gemini 2.5 Computer Use is Google's agentic model for browser automation, UI testing, and visual interface reasoning.

API: Yes Open weights: No Local: No
Provider
Google
Context
Unknown
Last verified
Needs recheck: verified June 24, 2026
Image, Multimodal, Text

Gemini 2.5 Flash

Gemini 2.5 Flash is a Google Gemini API model described as price-performance oriented for low-latency, high-volume tasks that require reasoning.

API: Yes Open weights: No Local: No
Provider
Google
Context
Unknown
Last verified
Needs recheck: verified June 18, 2026
Multimodal, Text

Gemini 2.5 Flash-Lite

Gemini 2.5 Flash-Lite is listed by Google as the fastest and most budget-friendly multimodal model in the Gemini 2.5 family.

API: Yes Open weights: No Local: No
Provider
Google
Context
Unknown
Last verified
Needs recheck: verified June 24, 2026
Image, Multimodal, Text

Gemini 3 Flash

Gemini 3 Flash is a Google Gemini API preview model positioned as frontier-class performance at a lower cost tier than larger models.

API: Yes Open weights: No Local: No
Provider
Google
Context
Unknown
Last verified
Needs recheck: verified June 18, 2026
Image, Multimodal, Text

Gemini 3.1 Flash-Lite

Gemini 3.1 Flash-Lite is a stable Google Gemini API model positioned for cost-efficient, high-volume agentic tasks, translation, and simpler data processing.

API: Yes Open weights: No Local: No
Provider
Google
Context
Unknown
Last verified
Needs recheck: verified June 18, 2026
Image, Multimodal, Text

Gemini 3.1 Pro

Gemini 3.1 Pro is a Google Gemini API preview model described for advanced intelligence, complex problem-solving, and agentic or vibe-coding capabilities.

API: Yes Open weights: No Local: No
Provider
Google
Context
Unknown
Last verified
Needs recheck: verified June 18, 2026
Image, Multimodal, Text

Gemini 3.5 Flash

Gemini 3.5 Flash is a Google Gemini API model listed as stable and positioned for sustained frontier performance on agentic and coding tasks.

API: Yes Open weights: No Local: No
Provider
Google
Context
Unknown
Last verified
Needs recheck: verified June 18, 2026
Image, Text

GPT-5.4

GPT-5.4 is an OpenAI API model described in official docs as a more affordable model for coding and professional work.

API: Yes Open weights: No Local: No
Provider
OpenAI
Context
1M tokens
Last verified
Needs recheck: verified June 18, 2026
Image, Text

GPT-5.4 mini

GPT-5.4 mini is an OpenAI API model described as a smaller model for coding, computer use, and subagents.

API: Yes Open weights: No Local: No
Provider
OpenAI
Context
400K tokens
Last verified
Needs recheck: verified June 18, 2026
Multimodal, Text

GPT-5.4 nano

GPT-5.4 nano is listed by OpenAI as the cheapest GPT-5.4-class model for simple high-volume tasks.

API: Yes Open weights: No Local: No
Provider
OpenAI
Context
1M tokens
Last verified
Needs recheck: verified June 24, 2026
Multimodal, Text

GPT-5.4 Pro

GPT-5.4 Pro is an OpenAI model catalog entry for teams that want a smarter GPT-5.4 variant for professional workflows.

API: Yes Open weights: No Local: No
Provider
OpenAI
Context
1M tokens
Last verified
Needs recheck: verified June 24, 2026
Image, Text

GPT-5.5

GPT-5.5 is OpenAI's flagship API model for complex reasoning, coding, and professional work, with text and image input and text output described in official OpenAI model docs.

API: Yes Open weights: No Local: No
Provider
OpenAI
Context
1M tokens
Last verified
Needs recheck: verified June 18, 2026
Multimodal, Text

GPT-5.5 Pro

GPT-5.5 Pro is an OpenAI frontier model variant listed in the OpenAI model catalog for higher-precision professional and coding work.

API: Yes Open weights: No Local: No
Provider
OpenAI
Context
1M tokens
Last verified
Needs recheck: verified June 24, 2026
Text

gpt-oss-20b

gpt-oss-20b is OpenAI's medium-sized open-weight gpt-oss model for low-latency, local, or specialized use cases.

API: No Open weights: Yes Local: Yes
Provider
OpenAI
Context
Unknown
Last verified
Needs recheck: verified June 24, 2026
Text

gpt-oss-120b

gpt-oss-120b is OpenAI's largest open-weight gpt-oss reasoning model, described by OpenAI as fitting into a single H100 GPU.

API: No Open weights: Yes Local: Yes
Provider
OpenAI
Context
Unknown
Last verified
Needs recheck: verified June 24, 2026
Image, Text

Grok 4.3

Grok 4.3 is an xAI API model listed with agentic tool calling, configurable reasoning, and a 1M-token context window.

API: Yes Open weights: No Local: No
Provider
xAI
Context
1M tokens
Last verified
Needs recheck: verified June 18, 2026
Text

Grok Build 0.1

Grok Build 0.1 is an xAI coding model described as trained specifically for fast agentic coding workflows.

API: Yes Open weights: No Local: No
Provider
xAI
Context
256K tokens
Last verified
Needs recheck: verified June 18, 2026
Text

Llama-3.1-Nemotron-Nano-8B-v1

Llama-3.1-Nemotron-Nano-8B-v1 is an NVIDIA compact open model that fits on a single RTX GPU and supports 128K context.

API: Yes Open weights: Yes Local: Yes
Provider
NVIDIA
Context
128K tokens
Last verified
Needs recheck: verified June 24, 2026
Text

Llama-3.3-Nemotron-Super-49B-v1.5

Llama-3.3-Nemotron-Super-49B-v1.5 is an NVIDIA reasoning model derived from Meta Llama 3.3 and tuned for RAG and tool calling.

API: Yes Open weights: Yes Local: Yes
Provider
NVIDIA
Context
128K tokens
Last verified
Needs recheck: verified June 24, 2026
Image, Multimodal, Text

Llama 4 Maverick

Llama 4 Maverick is a larger Meta Llama 4 mixture-of-experts model optimized for multimodal understanding, coding, tool-calling, and agentic systems.

API: Unknown Open weights: Yes Local: Yes
Provider
Meta
Context
1M tokens
Last verified
Needs recheck: verified June 18, 2026
Image, Multimodal, Text

Llama 4 Scout

Llama 4 Scout is a Meta Llama 4 mixture-of-experts model optimized for multimodal understanding, multilingual tasks, coding, tool-calling, and agentic systems.

API: Unknown Open weights: Yes Local: Yes
Provider
Meta
Context
10M tokens
Last verified
Needs recheck: verified June 18, 2026
Text

Magistral Medium 1.2

Magistral Medium 1.2 is a Mistral model listed in Mistral documentation for reasoning-oriented workflows.

API: Yes Open weights: No Local: No
Provider
Mistral AI
Context
128k tokens
Last verified
Needs recheck: verified June 24, 2026
Text

Ministral 3 14B

Ministral 3 14B is a Mistral small-model variant listed with the Ministral 3 family in official documentation.

API: Yes Open weights: Yes Local: Yes
Provider
Mistral AI
Context
256k tokens
Last verified
Needs recheck: verified June 24, 2026