DeepSeek V4 'Full-Blooded Edition' Leaked, Could Be Released As Early As Tomorrow

marsbitPublished on 2026-07-19Last updated on 2026-07-19

Abstract

The highly anticipated full release of DeepSeek V4 is imminent, expected to launch as early as tomorrow after nearly three months of waiting. A select group has already received access to the GA (General Availability) beta, which includes two versions: DeepSeek V4 Flash and DeepSeek V4 Pro. Early testers report that V4's overall performance is close to the level of Opus 4.8, with coding capabilities rivaling GPT-5.6 Sol. Its agent abilities are significantly enhanced, and 3D/SVG generation has improved notably. While it may not surpass the recently released Kimi K3 in performance, its expected price point is significantly lower. The official release will introduce a new "peak/off-peak" pricing model for its API. For example, deepseek-v4-pro will cost $0.87 per million output tokens during standard times and $1.74 during peak hours. The flash version is even more aggressive at $0.28/$0.56 per million tokens, with cached input tokens priced extremely low at $0.0028. This makes V4 a strong contender in terms of cost-effectiveness, potentially offering Opus-level capabilities at a fraction of the cost, continuing DeepSeek's reputation as a "price disruptor" in the AI market. Initial demos showcasing V4's capabilities have begun circulating, including generated 3D simulation games, HTML games blending elements of Minecraft and No Man's Sky, and classic games like a "Cut the Rope" clone. The final GA version is set to replace the older deepseek-chat and deepseek-reasoner models,...

The whole internet has been waiting for nearly three months!

DeepSeek V4 official version may be released as early as tomorrow, at the latest within the next few days.

Currently, some people have already received early access to the DeepSeek V4 (GA) grayscale testing.

There are two versions: DeepSeek V4 Flash and DeepSeek V4 Pro.

DeepSeek V4 'Full-Blooded Edition'

Is Finally Coming

What everyone is most concerned about is, how do you know if you've been included in the grayscale test?

Blogger AiBattle gave a folk 'verification mantra': look at the first person in the Chain-of-Thought (CoT).

If the model's thought process starts with "I'm" or "I'll", not the old version's "Let me".

Then congratulations, you're probably already using the V4 GA version!

In the current landscape where Fable 5 and GPT-5.6 Sol are locked in a fierce battle, the entire internet's expectations for the open-source AI's big move are sky-high.

A developer, Pankaj Kumar, after experiencing it, first gave a fair summary of V4's performance:

Overall performance approaches Opus 4.8 level, coding rivals GPT-5.6 Sol;

Agent capabilities are greatly enhanced, 3D and SVG generation have significantly improved;

For the same task, V4 requires more iteration rounds than Fable 5.

He stated that from the current market positioning, V4 will likely not beat the newly released Kimi K3, but the price will be significantly lower.

If this performance truly comes with this price, then it might be another DeepSeek moment.

First Round of Tests Already Out

Now, the first round of test demos for DeepSeek V4 (GA) have begun to circulate.

It can be said that the outside world has mixed reviews:

Some believe it can already rival Claude 5, but some developers also pointed out that the Pro version is not significantly ahead of the Flash version.

Below is a 3D simulated shooting game generated by V4 Pro. The core gameplay involves operating a ballista for target practice, and basic UI functions are already complete.

An HTML hybrid game blending Minecraft and No Man's Sky, created by the V4 official version, has relatively high playability.

The classic "Cut the Rope" game below was also produced by V4 in one go.

Below are also some demos generated by V4: an Xbox controller SVG test and game generation.

Introducing "Peak-Valley Billing" for the First Time

Still Smells Sweet

To recreate a "DeepSeek moment," the biggest variable is price.

All leaks point to the same conclusion: performance might not be number one, but the price will be significantly lower.

At the end of last month, DeepSeek sent an email to all API users, stating that the V4 official version would be launched in mid-July.

After the GA release, API pricing will be adjusted simultaneously, introducing "Peak-Valley Billing."

deepseek-v4-pro: $0.87 per million output tokens during off-peak, $1.74 during peak; $0.435 per million for cache-missed inputs during off-peak.

deepseek-v4-flash is even more aggressive: $0.28 per million outputs during off-peak, $0.56 during peak, $0.0028 per million for cache-hit inputs.

This means the "price butcher" who never raised prices has, for the first time, installed a meter on its computing power.

The impact on individual users is minimal, but for teams that run Agents non-stop during working hours, this is a real cost that needs to be recalculated.

Fortunately, the price for cache hits remains extremely low. Moving batch tasks, evaluation benchmarks, and data generation outside peak hours can still bring costs back down.

However, compared to Fable 5's $50 per million output tokens, V4 remains the most cost-effective option.

After all, the performance has also reached the Opus level. In the official self-test on SWE-bench Verified when the V4 preview was released—

DeepSeek-V4-Pro-Max was only 0.2 percentage points behind Claude Opus 4.6 Max, yet the price was only one-seventh of the latter's.

But in long-context retrieval, professional software engineering, and some knowledge reasoning evaluations, Opus still maintains a lead.

As a side note, the two old model names, deepseek-chat and deepseek-reasoner, will be officially discontinued on July 24th.

Regardless, this "shoe" that the whole internet has been waiting three months for is finally about to drop.

Judging from the currently leaked information, V4 will likely not be the model that is "number one in all categories."

With Fable 5 and GPT-5.6 Sol ahead, and Kimi K3 on the side, it's difficult for V4 to overturn the table relying solely on pure performance.

But DeepSeek has never played that card.

The real highlight is that old, repeatedly proven path: offering Opus-level capabilities at one-seventh the price.

After all, in the face of giants charging dozens of dollars per million tokens, even with a "peak-valley meter" installed, V4 is still that "price butcher" that wreaks havoc.

References:

https://x.com/pankajkumar_dev/status/2078536231026372846

This article is from the WeChat public account "New Zhiyuan," edited by Peach.

Related Questions

QWhat are the two versions of DeepSeek V4 mentioned in the article?

AThe two versions mentioned are DeepSeek V4 Flash and DeepSeek V4 Pro.

QHow can users check if they have access to the DeepSeek V4 GA (General Availability) version, according to the blogger AiBattle?

AAccording to blogger AiBattle, users can check the first-person pronoun in the model's Chain-of-Thought (CoT). If the thought process begins with "I'm" or "I'll" instead of the older version's "Let me," they are likely using the V4 GA version.

QWhat is a key pricing feature that DeepSeek is introducing with the V4 GA release?

ADeepSeek is introducing a 'peak and off-peak billing' (or time-of-use pricing) feature with the V4 GA release, where API costs vary based on usage during peak or off-peak times.

QWhat is the estimated performance level of DeepSeek V4 according to developer Pankaj Kumar's initial assessment?

AAccording to developer Pankaj Kumar's assessment, DeepSeek V4's overall performance is close to the level of Opus 4.8, with coding capabilities rivaling GPT-5.6 Sol, but it may not surpass the recently released Kimi K3.

QWhen are the older model names 'deepseek-chat' and 'deepseek-reasoner' scheduled to be officially discontinued?

AThe older model names 'deepseek-chat' and 'deepseek-reasoner' are scheduled to be officially discontinued on July 24th.

Related Reads

Altman's Bold Statement: In Six Months, GPT's Context Will Cover Everything About You

OpenAI CEO Sam Altman, in a recent interview, stated that the next generation of ChatGPT, expected within six months, will evolve into an active workplace collaborator. The envisioned model will be able to monitor your screen, record meetings and calls, and integrate with emails, messages, and documents—essentially gaining a complete understanding of your work and life context. The core goal is to combat information overload by proactively identifying missed details, summarizing key points from feedback, and highlighting critical information from vast data streams, rather than merely responding to queries. Altman specifically targets startup CEOs as the primary users, acknowledging their struggle with overwhelming, never-ending tasks. The AI would act as a 24/7 assistant that understands context, helps draft documents, reminds users of client concerns, and fills gaps in strategic plans—all while leaving final decision-making to humans. The announcement sparked a polarized public reaction. Supporters welcome the potential relief from administrative burdens, while critics express significant privacy concerns, fearing extensive data collection. Skeptics also question whether such deep integration into personal and professional life is necessary or safe. Altman emphasizes that user permission and control over data access are fundamental, but the central challenge remains: balancing vastly improved AI utility with the privacy trade-offs required for such deep contextual awareness.

marsbit3m ago

Altman's Bold Statement: In Six Months, GPT's Context Will Cover Everything About You

marsbit3m ago

The Most Mysterious Boss Lady: She Founded a VC Firm

"The Most Mysterious 'Boss's Wife' Who Started a VC" Recently, a name has been frequently mentioned among investors—Yang Jingting, the founder behind XINHE Capital. While not widely known herself, her husband is Liu Sheng, Chairman of trillion-dollar market cap company Zhongji Innolight. The couple entered the public eye during a Tsinghua University donation ceremony in April. However, Yang has been quietly building a significant investment portfolio through XINHE Capital since 2021, where she holds a 90% stake. With a background starting in 2014 across various investment firms, she now focuses XINHE Capital on upstream photonics and semiconductor sectors like materials, chips, and equipment, targeting data centers, smart devices, and automotive. The fund has accelerated its pace in 2024. Yang's broader activities include direct investments (e.g., Suzhou Suna Optoelectronics) and acting as an LP in funds like those from Yongxin Fangzhou and Qianrong Holdings. Her investments are strategically separate from her husband's core business. Her story is part of a larger trend: China's industrial tycoons are increasingly moving into tech VC. Examples include "Glass Queen" Zhou Qunfei (Lens Technology) investing in AI, Zhu Xingming (Inovance Technology) via family office Minghui Investment targeting brain-computer interfaces, and Luxshare Precision's family office. Liu Yi (Andon Health) recently participated in DeepSeek's funding round. These entrepreneurs leverage their industrial expertise and capital to back next-generation technologies like AI, semiconductors, and embodied intelligence, shifting the flow of wealth from real estate towards defining the future of hard tech.

marsbit9m ago

The Most Mysterious Boss Lady: She Founded a VC Firm

marsbit9m ago

Trading

Spot
活动图片