Kingy AI Launch Intelligence
AI Model Intelligence Hub
Kingy AI tracks AI models by provider, model family, modality, release status, API access, web access, local availability, open-weight/open-source status, pricing notes, hardware requirements, official sources, benchmark caveats, and last-verified status.
Source-checked profiles within the review window appear first; retired models appear last. Check the date and limitations on each record. Indexing is assessed separately from review age.
Image, Multimodal, Text
Google’s stable Gemini 3.1 Flash-Lite accepts text, images, video, audio and PDFs and produces text. This is a documentation review, not a Kingy performance test.
API: Yes Open weights: Unknown Local: Unknown
- Provider
- Google
- Context
- Input token limit: 1,048,576
- Last verified
- 2026-09-14
Audio, Multimodal, Video
MiniMax H3 is a multimodal video-generation model with distinct hosted API, Hailuo product, and open-weight access routes. This living record keeps those access paths and their evidence boundaries separate.
API: Yes Open weights: Yes Local: Yes
- Provider
- MiniMax Group
- Context
- Unknown
- Last verified
- 2026-08-25
Image, Multimodal, Text
Gemini 2.5 Flash is a Google Gemini API model described as price-performance oriented for low-latency, high-volume tasks that require reasoning.
API: Yes Open weights: No Local: No
- Provider
- Google
- Context
- Unknown
- Last verified
- 2026-08-16
Multimodal, Text
Gemini 2.5 Flash-Lite is listed by Google as the fastest and most budget-friendly multimodal model in the Gemini 2.5 family.
API: Yes Open weights: No Local: No
- Provider
- Google
- Context
- Unknown
- Last verified
- 2026-08-16
Image, Multimodal
Gemini 2.5 Flash Image, also known as Nano Banana, is Google's image generation and editing model for high-volume visual workflows.
API: Yes Open weights: No Local: No
- Provider
- Google
- Context
- N/A for image generation
- Last verified
- 2026-08-16
Image, Multimodal, Text
Gemini 2.5 Pro is a Google Gemini API model described as advanced for complex tasks, deep reasoning, and coding capabilities.
API: Yes Open weights: No Local: No
- Provider
- Google
- Context
- Unknown
- Last verified
- 2026-08-16
Multimodal, Text
Claude Fable 5 is Anthropic's most capable widely released model, listed for demanding reasoning and long-horizon agentic work.
API: Yes Open weights: No Local: No
- Provider
- Anthropic
- Context
- 1M tokens
- Last verified
- Needs recheck: checked July 28, 2026
Text
DeepSeek V4 Pro is listed in DeepSeek API docs as a current model supporting thinking and non-thinking modes, JSON output, tool calls, and a 1M context length.
API: Yes Open weights: No Local: No
- Provider
- DeepSeek
- Context
- 1M tokens
- Last verified
- Needs recheck: checked July 28, 2026
Image, Multimodal, Text
Gemini 3.6 Flash is Google's latest Flash-tier model, balancing speed with intelligence for agentic and multimodal tasks.
API: Yes Open weights: No Local: No
- Provider
- Google
- Context
- 1,048,576 tokens
- Last verified
- Needs recheck: checked July 28, 2026
Image, Multimodal, Text
Mistral Medium 3.5 is listed by Mistral as a frontier-class multimodal model optimized for agentic and coding use cases.
API: Yes Open weights: Yes Local: Yes
- Provider
- Mistral AI
- Context
- 256K tokens
- Last verified
- Needs recheck: checked July 28, 2026
Video
Kling Video 3.0 is Kuaishou’s February 2026 short-form video model for text-, image- and reference-guided generation, with multi-shot control, element consistency, native multilingual audio and flexible 3–15 second output.
API: Yes Open weights: No Local: No
- Provider
- Kling AI / Kuaishou Technology
- Context
- Not applicable to video generation
- Last verified
- Needs recheck: checked July 27, 2026
Multimodal, Text
Claude Opus 5 is Anthropic's recommended default model for complex agentic coding and enterprise work, with a 1M-token context window and 128K maximum output.
API: Yes Open weights: No Local: No
- Provider
- Anthropic
- Context
- 1M tokens
- Last verified
- Needs recheck: checked July 25, 2026
Multimodal, Text
Claude Sonnet 5 is Anthropic's speed-and-intelligence balance in the Claude 5 line, with a 1M-token context window and 128K maximum output.
API: Yes Open weights: No Local: No
- Provider
- Anthropic
- Context
- 1M tokens
- Last verified
- Needs recheck: checked July 25, 2026
Image, Text
GPT-5.6 Luna is the cost-optimised tier of the GPT-5.6 family, retaining the full 1,050,000-token context window and reasoning-token support.
API: Yes Open weights: No Local: No
- Provider
- OpenAI
- Context
- 1,050,000 tokens
- Last verified
- Needs recheck: checked July 25, 2026
Image, Text
GPT-5.6 Sol is OpenAI's frontier model for complex professional work, and the model the bare gpt-5.6 alias resolves to.
API: Yes Open weights: No Local: No
- Provider
- OpenAI
- Context
- 1,050,000 tokens
- Last verified
- Needs recheck: checked July 25, 2026
Image, Text
GPT-5.6 Terra is the balanced intelligence-and-cost tier of the GPT-5.6 family, positioned roughly where the mini tier sat in earlier GPT-5 generations.
API: Yes Open weights: No Local: No
- Provider
- OpenAI
- Context
- 1,050,000 tokens
- Last verified
- Needs recheck: checked July 25, 2026
Multimodal, Text
Claude Mythos 5 is an Anthropic Project Glasswing model listed as limited availability for approved customers.
API: Yes Open weights: No Local: No
- Provider
- Anthropic
- Context
- 1M tokens
- Last verified
- Needs recheck: checked June 24, 2026
Code, Text
Codestral is Mistral AI's coding-focused model line, listed in Mistral model documentation and context-window notes.
API: Yes Open weights: No Local: No
- Provider
- Mistral AI
- Context
- 128k tokens
- Last verified
- Needs recheck: checked June 24, 2026
Text
Command A is Cohere's enterprise model for tool use, RAG, agents, and multilingual work, listed with 111B parameters and 256K context.
API: Yes Open weights: No Local: No
- Provider
- Cohere
- Context
- 256K tokens
- Last verified
- Needs recheck: checked June 24, 2026
Multimodal, Text
Command A+ is Cohere's Command A family model announced with vision inputs, reasoning, translation, and agentic task support.
API: Yes Open weights: No Local: No
- Provider
- Cohere
- Context
- Unknown
- Last verified
- Needs recheck: checked June 24, 2026
Text
Command R7B is Cohere's smallest and fastest R-family model for RAG, tool use, and agents.
API: Yes Open weights: No Local: No
- Provider
- Cohere
- Context
- 128K tokens
- Last verified
- Needs recheck: checked June 24, 2026
Text
Command R+ is Cohere's R-family model optimized for conversational interaction, long-context tasks, complex RAG, and multi-step tool use.
API: Yes Open weights: No Local: No
- Provider
- Cohere
- Context
- 128K tokens
- Last verified
- Needs recheck: checked June 24, 2026
Embeddings, Image, Text
Embed multilingual v3.0 is a Cohere embedding model for multilingual classification and embedding support.
API: Yes Open weights: No Local: No
- Provider
- Cohere
- Context
- 512 token embedding input in Cohere docs table
- Last verified
- Needs recheck: checked June 24, 2026
Image
FLUX.1 Kontext [pro] is a BFL model for text-to-image generation and context-aware image editing.
API: Yes Open weights: No Local: No
- Provider
- Black Forest Labs
- Context
- N/A for image generation
- Last verified
- Needs recheck: checked June 24, 2026