AI Model Profile

DeepSeek V4 Pro

DeepSeek V4 Pro is listed in DeepSeek API docs as a current model supporting thinking and non-thinking modes, JSON output, tool calls, and a 1M context length.

Family
DeepSeek V4
Release date
Unknown
Status
Current
Context window
1M tokens
Output limit
384K tokens maximum
API
yes
Open weights
no
Local/self-hosted
no
Pricing
DeepSeek API pricing checked 2026-07-28: $0.435 per million cache-miss input tokens, $0.003625 per million cache-hit input tokens and $0.87 per million output tokens.
Verification
verified

Verification & Sources

Status
Verified
Source links
5
Freshness
Verified July 28, 2026
Last verified
July 28, 2026
Last updated
July 28, 2026
Suggest a correction

Form submissions, correction notes, score details, URLs, and analytics events may be stored for editorial review, spam prevention, product improvement, and follow-up. Do not submit secrets, unreleased financials, private customer data, or regulated personal data through these forms.

Benchmark Caveat

Provider documentation and provider-published capability notes are directional, not universal rankings. Real results depend on prompts, tools, latency targets, pricing tier, safety filters, context length, and workload mix.

Best for

Candidate for DeepSeek API workflows that need the Pro variant and long-context reasoning or coding tests.

Skip if

Skip if your team needs independently benchmarked performance claims or contractual guarantees not present in the official provider documentation.

Strengths

Official docs list a long context window, tool calls, JSON output, and current V4 Pro access.

Weaknesses

This profile does not independently validate benchmark quality, reliability, or latency.

Agent suitability

Candidate for tool-using agents based on official API feature support.

Kingy AI take

Use this profile as a source-backed reference point, not a ranking. Re-check official provider docs before making production or budget decisions.

Full Model Notes

DeepSeek V4 Pro is listed in DeepSeek API docs as a current model supporting thinking and non-thinking modes, JSON output, tool calls, and a 1M context length.

Coding notes

Evaluate coding workflows directly, especially for agentic code assistants.

Reasoning notes

Official docs describe thinking and non-thinking mode support.

Creative notes

Not primarily a creative/media model in this profile.

Research notes

Candidate for long-context research tests with current API docs.

API pricing notes

Record returned model, system_fingerprint and usage. Thinking is the default. Publication requires clear AI-output disclosure; DeepSeek privacy materials describe processing and storage in the People's Republic of China.

License notes

Review the official provider terms and model documentation before relying on license or redistribution assumptions.