DeepSeekPremium plan

DeepSeek V4 Pro

DeepSeek-V4-Pro-0813 via Fireworks: official DeepSeek V4 Pro release with stronger agentic performance, served as a serverless 1.6T MoE at 1M context scale. Function calling supported. Uses 1 premium request per send before length multipliers. Does not support web search or image input.

You always get the exact model you pick — we never silently route you to another.

About DeepSeek V4 Pro

DeepSeek V4 Pro makes a compelling case that frontier-class coding performance and a one-million-token context window do not have to cost frontier-class money. The official 0813 release supersedes the preview, adding a DSpark speculative-decoding module and stronger agentic performance in production. On just4o.chat it runs via Fireworks serverless as deepseek-v4-pro-0813. Users consistently praise its agentic coding ability, noting it competes with or beats larger closed models on multi-step coding tasks, and its hybrid attention architecture handles full-codebase analysis without collapsing under the token budget. The MIT license is a genuine differentiator: weights are freely available for self-hosting, fine-tuning, and commercial integration. The honest caveat: V4 Pro is verbose. It can generate four to five times more output tokens than comparable models on the same prompt, which erodes the per-token savings and makes cost estimation harder than it first appears.

Best for

  • High-volume automated coding pipelines, code review, and refactoring where per-call cost matters
  • Full-codebase analysis using 1M-token context for migration planning or architectural review
  • Multi-turn agentic workflows where reasoning must persist across tool calls and conversation turns
  • Long-document synthesis, research corpora summarization, and extraction from large structured data
  • On-premises or self-hosted deployment via MIT-licensed weights with custom fine-tuning

Specifications

ProviderDeepSeek
Released2026-08-13
IntelligenceHigh
SpeedMedium
Context window1,048,600 tokens
Max output384,000 tokens
Knowledge cutoffAugust 2026
Input price$1.32 / 1M tokens
Output price$3.96 / 1M tokens
Request cost1 premium request
Plan tierPremium
InputText
OutputText
FeaturesCached input: $0.044 / 1M tokens, 1 premium request per send before length multipliers, Function calling supported, Serverless through Fireworks (deepseek-v4-pro-0813), On-demand Fireworks deployments available separately
Model IDdeepseek-v4-pro

Frequently asked questions

On Fireworks serverless, input tokens are $1.32 per million, cached input is $0.044 per million, and output tokens are $3.96 per million.

1,048,576 tokens (approximately 1 million tokens), with a maximum output of 384,000 tokens.

Not yet as of mid-2026. All published scores — including 80.6% SWE-Bench Verified and 90.1% GPQA Diamond — come from DeepSeek's internal evaluations. Independent third-party leaderboard verification is pending.

It produces significantly more output tokens than most models on the same prompts — sometimes 4 to 5 times more — making real costs harder to predict. It also has a reliability gap versus top closed models on the most complex edge-case reasoning tasks, and it censors politically sensitive topics related to Chinese governance.

Teams running cost-sensitive, high-volume coding or document workloads who can self-host or accept the API's privacy trade-offs, and who want MIT-licensed weights for fine-tuning or commercial integration.

DeepSeek V4 Pro maps to accounts/fireworks/models/deepseek-v4-pro-0813, the official August 13 release that supersedes the preview serverless endpoint.

The 0813 build is the official DeepSeek V4 Pro release and supersedes the preview version.

Related models