DeepSeek V4 'Full-Blooded Edition' Leaked, Could Be Released As Early As Tomorrow

marsbitPublicado em 2026-07-19Última atualização em 2026-07-19

Resumo

The highly anticipated full release of DeepSeek V4 is imminent, expected to launch as early as tomorrow after nearly three months of waiting. A select group has already received access to the GA (General Availability) beta, which includes two versions: DeepSeek V4 Flash and DeepSeek V4 Pro. Early testers report that V4's overall performance is close to the level of Opus 4.8, with coding capabilities rivaling GPT-5.6 Sol. Its agent abilities are significantly enhanced, and 3D/SVG generation has improved notably. While it may not surpass the recently released Kimi K3 in performance, its expected price point is significantly lower. The official release will introduce a new "peak/off-peak" pricing model for its API. For example, deepseek-v4-pro will cost $0.87 per million output tokens during standard times and $1.74 during peak hours. The flash version is even more aggressive at $0.28/$0.56 per million tokens, with cached input tokens priced extremely low at $0.0028. This makes V4 a strong contender in terms of cost-effectiveness, potentially offering Opus-level capabilities at a fraction of the cost, continuing DeepSeek's reputation as a "price disruptor" in the AI market. Initial demos showcasing V4's capabilities have begun circulating, including generated 3D simulation games, HTML games blending elements of Minecraft and No Man's Sky, and classic games like a "Cut the Rope" clone. The final GA version is set to replace the older deepseek-chat and deepseek-reasoner models,...

The whole internet has been waiting for nearly three months!

DeepSeek V4 official version may be released as early as tomorrow, at the latest within the next few days.

Currently, some people have already received early access to the DeepSeek V4 (GA) grayscale testing.

There are two versions: DeepSeek V4 Flash and DeepSeek V4 Pro.

DeepSeek V4 'Full-Blooded Edition'

Is Finally Coming

What everyone is most concerned about is, how do you know if you've been included in the grayscale test?

Blogger AiBattle gave a folk 'verification mantra': look at the first person in the Chain-of-Thought (CoT).

If the model's thought process starts with "I'm" or "I'll", not the old version's "Let me".

Then congratulations, you're probably already using the V4 GA version!

In the current landscape where Fable 5 and GPT-5.6 Sol are locked in a fierce battle, the entire internet's expectations for the open-source AI's big move are sky-high.

A developer, Pankaj Kumar, after experiencing it, first gave a fair summary of V4's performance:

Overall performance approaches Opus 4.8 level, coding rivals GPT-5.6 Sol;

Agent capabilities are greatly enhanced, 3D and SVG generation have significantly improved;

For the same task, V4 requires more iteration rounds than Fable 5.

He stated that from the current market positioning, V4 will likely not beat the newly released Kimi K3, but the price will be significantly lower.

If this performance truly comes with this price, then it might be another DeepSeek moment.

First Round of Tests Already Out

Now, the first round of test demos for DeepSeek V4 (GA) have begun to circulate.

It can be said that the outside world has mixed reviews:

Some believe it can already rival Claude 5, but some developers also pointed out that the Pro version is not significantly ahead of the Flash version.

Below is a 3D simulated shooting game generated by V4 Pro. The core gameplay involves operating a ballista for target practice, and basic UI functions are already complete.

An HTML hybrid game blending Minecraft and No Man's Sky, created by the V4 official version, has relatively high playability.

The classic "Cut the Rope" game below was also produced by V4 in one go.

Below are also some demos generated by V4: an Xbox controller SVG test and game generation.

Introducing "Peak-Valley Billing" for the First Time

Still Smells Sweet

To recreate a "DeepSeek moment," the biggest variable is price.

All leaks point to the same conclusion: performance might not be number one, but the price will be significantly lower.

At the end of last month, DeepSeek sent an email to all API users, stating that the V4 official version would be launched in mid-July.

After the GA release, API pricing will be adjusted simultaneously, introducing "Peak-Valley Billing."

deepseek-v4-pro: $0.87 per million output tokens during off-peak, $1.74 during peak; $0.435 per million for cache-missed inputs during off-peak.

deepseek-v4-flash is even more aggressive: $0.28 per million outputs during off-peak, $0.56 during peak, $0.0028 per million for cache-hit inputs.

This means the "price butcher" who never raised prices has, for the first time, installed a meter on its computing power.

The impact on individual users is minimal, but for teams that run Agents non-stop during working hours, this is a real cost that needs to be recalculated.

Fortunately, the price for cache hits remains extremely low. Moving batch tasks, evaluation benchmarks, and data generation outside peak hours can still bring costs back down.

However, compared to Fable 5's $50 per million output tokens, V4 remains the most cost-effective option.

After all, the performance has also reached the Opus level. In the official self-test on SWE-bench Verified when the V4 preview was released—

DeepSeek-V4-Pro-Max was only 0.2 percentage points behind Claude Opus 4.6 Max, yet the price was only one-seventh of the latter's.

But in long-context retrieval, professional software engineering, and some knowledge reasoning evaluations, Opus still maintains a lead.

As a side note, the two old model names, deepseek-chat and deepseek-reasoner, will be officially discontinued on July 24th.

Regardless, this "shoe" that the whole internet has been waiting three months for is finally about to drop.

Judging from the currently leaked information, V4 will likely not be the model that is "number one in all categories."

With Fable 5 and GPT-5.6 Sol ahead, and Kimi K3 on the side, it's difficult for V4 to overturn the table relying solely on pure performance.

But DeepSeek has never played that card.

The real highlight is that old, repeatedly proven path: offering Opus-level capabilities at one-seventh the price.

After all, in the face of giants charging dozens of dollars per million tokens, even with a "peak-valley meter" installed, V4 is still that "price butcher" that wreaks havoc.

References:

https://x.com/pankajkumar_dev/status/2078536231026372846

This article is from the WeChat public account "New Zhiyuan," edited by Peach.

Perguntas relacionadas

QWhat are the two versions of DeepSeek V4 mentioned in the article?

AThe two versions mentioned are DeepSeek V4 Flash and DeepSeek V4 Pro.

QHow can users check if they have access to the DeepSeek V4 GA (General Availability) version, according to the blogger AiBattle?

AAccording to blogger AiBattle, users can check the first-person pronoun in the model's Chain-of-Thought (CoT). If the thought process begins with "I'm" or "I'll" instead of the older version's "Let me," they are likely using the V4 GA version.

QWhat is a key pricing feature that DeepSeek is introducing with the V4 GA release?

ADeepSeek is introducing a 'peak and off-peak billing' (or time-of-use pricing) feature with the V4 GA release, where API costs vary based on usage during peak or off-peak times.

QWhat is the estimated performance level of DeepSeek V4 according to developer Pankaj Kumar's initial assessment?

AAccording to developer Pankaj Kumar's assessment, DeepSeek V4's overall performance is close to the level of Opus 4.8, with coding capabilities rivaling GPT-5.6 Sol, but it may not surpass the recently released Kimi K3.

QWhen are the older model names 'deepseek-chat' and 'deepseek-reasoner' scheduled to be officially discontinued?

AThe older model names 'deepseek-chat' and 'deepseek-reasoner' are scheduled to be officially discontinued on July 24th.

Leituras Relacionadas

The Once-Niche Field of Philosophy Becomes a Hot Topic in "Governing" AI

"Cold" Philosophy Becomes Hot in Taming AI This article explores the rising prominence of philosophical inquiry at the World Artificial Intelligence Conference (WAIC), highlighting a shift from purely technical discussions to deeper questions about AI's nature and impact. A key theme is the foundational question of intelligence itself. Philosopher Sun Ning argued that true, grounded intelligence requires embodiment, environment, interaction with others, and historical context, summarized as "Before intelligence, there is a world. Before mind, there is relationship." The forum then examined AI's expanding role in science. Researchers presented AI systems that can autonomously generate scientific papers and mathematical proofs, raising critical questions about evaluation, attribution of discovery, and potential misuse. The proposed solution is a collaborative framework: AI expands the search space, humans define value and provide rigorous constraints, and machines handle verification. As AI models begin to simulate human societies and behaviors for research, new risks emerge. Simulations can inherit and amplify societal biases from their training data, and their outputs risk being mistaken for genuine social signals. This necessitates robust governance focused on auditability, bias correction, and clear human oversight—ensuring people retain intervention, correction, and explanation rights ("human-in-the-loop"). The discussion extended to industry, where AI integrates with wet labs for bio-manufacturing. Here, challenges involve bridging the digital-physical gap and adapting regulatory and business models from process-based to outcome-based partnerships, fundamentally altering production relationships. The convergence of philosophers, scientists, and entrepreneurs at WAIC signals a broader trend: as AI permeates knowledge creation and social systems, technical progress is increasingly intertwined with urgent questions of ethics, governance, economic distribution, and public policy. The core challenge becomes not just what AI *can* do, but what we *should* ask it, trust from it, and guide it towards.

marsbitHá 51m

The Once-Niche Field of Philosophy Becomes a Hot Topic in "Governing" AI

marsbitHá 51m

After six years of pie-in-the-sky fundraising, 90% of funds used to replenish cash flow: The illusion and reality of L'CI Technology's silicon carbide "industrialization"|TMTPost Deep Dive

After six years and raising approximately 3.2 billion yuan through two private placements, Luxshare Technology has officially terminated its two major SiC wafer projects. Only about 7.75% of the raised funds were actually invested in SiC construction and R&D, with the remaining 90% redirected to replenish working capital and repay loans. Despite previous ambitious plans for a 10-billion-yuan SiC industrial park and public assurances of progress, the company's actual SiC capacity remains unclear, with its subsidiary recording significant losses. The article details a pattern of chasing market trends, from electric vehicles and photovoltaics to the current focus on SiC. However, its SiC entry was late; while peers like Tianke Heda and Tianyue Advanced have moved from mass-producing 6-inch to 8-inch and even 12-inch wafers, Luxshare is still struggling with its 6-inch plans and now promises to develop larger sizes with its own funds. Key concerns include unclear disclosures about the status of its chief scientist, Chen Zhizhan, and significantly lower R&D investment compared to competitors. Financially, since its 2011 IPO, the company has raised over 7 billion yuan but generated only 221 million yuan in cumulative net profit. Meanwhile, the controlling shareholder family has cashed out nearly 2 billion yuan since 2018. The article positions Luxshare as a case study of a listed company focused more on financing and capital operation around hot topics than on substantive industrial development.

marsbitHá 1h

After six years of pie-in-the-sky fundraising, 90% of funds used to replenish cash flow: The illusion and reality of L'CI Technology's silicon carbide "industrialization"|TMTPost Deep Dive

marsbitHá 1h

ChatGPT Finally Can 'Search Itself': Nearly Four Years of Conversations, Retrieved with One Click

ChatGPT, nearly four years old, has long lacked a functional way for users to search their own accumulated history. This changed on July 14th, when OpenAI launched a comprehensive "Unified Search" feature across all platforms and plans. This new search allows users to find and instantly access past conversations, uploaded documents, generated images, and projects within their own ChatGPT account, all from a single sidebar entry. This update signifies a major shift in ChatGPT's role, transforming it from a transient chat tool into a personal knowledge base or "file system" for users' AI-generated content. It completes a critical missing piece as OpenAI expands ChatGPT's capabilities with features like Projects, Memory, Work, and Sites, aiming to make it a central hub for work and creativity. The ability to easily retrieve years of personal data makes that data more valuable as a unique, irreplaceable asset in the AI era. However, it also highlights the increasing weight and potential sensitivity of this data, given default settings that allow conversations to be used for model training and legal precedents involving user logs. While a significant improvement over the previous ineffective search, this move is seen as addressing an industry-wide oversight. As the competition between AI assistants (ChatGPT, Gemini, Claude) moves beyond raw model power, the new battleground is becoming which platform can best organize, retain, and leverage the valuable data co-created by users and AI.

marsbitHá 1h

ChatGPT Finally Can 'Search Itself': Nearly Four Years of Conversations, Retrieved with One Click

marsbitHá 1h

Trading

Spot
活动图片