Cook Fires Off His Last Shots: 2nm, Tao's Law, the Most Powerful AI Computer

marsbit2026-08-26 tarihinde yayınlandı2026-08-26 tarihinde güncellendi

Özet

Apple's CEO Tim Cook unveiled the company's final major chip updates under his tenure, introducing the first 2nm chip M6 and the powerful M5 Ultra with a quad-die architecture. The M6, featured in the new Mac mini, offers up to 4x faster AI performance and a starting price of 6,999 RMB, a significant increase from its predecessor. The M5 Ultra powers the Mac Studio with up to 36-core CPU, 80-core GPU, and a massive 512GB of unified memory, delivering up to 4.3x faster AI performance than the M3 Ultra. Both chips emphasize enhanced on-device AI capabilities, positioning Macs for agentic workflows and local large language model inference. However, high demand for memory in the AI sector has led to supply constraints, with the 512GB Mac Studio model delayed until late October. These releases mark a strategic push into high-performance AI computing for Apple's desktop lineup as Cook prepares to step down.

Old man Cook's final dance before retirement, he really emptied the magazine and went all out!!!

The first 2nm chip, the M6, is here.

But the new-process M6 isn't the most powerful; that title belongs to the newly released M5 Ultra, arguably the strongest M-series chip in Apple's history.

Although it uses a 3nm process, Apple specifically emphasized using a pioneering 4-die stacking technology—

This can be considered an application of Huawei's proposed Tao's Law.

Not only that, these two chips are directly packed into the new Macs—

The Mac mini gets the M6/M5 Pro, with a China retail price starting at 6999 yuan; the Mac Studio is upgraded with the M5 Max/M5 Ultra, starting at 19999 yuan in China.

The chassis remains largely familiar, but inside, it's charging full speed ahead into AI.

The AI performance of the M6 Mac mini surges up to 4 times.

The M5 Ultra Mac Studio features up to a 36-core CPU, 80-core GPU, and can be configured with up to 512GB of unified memory.

More interestingly, as soon as this wave of AI computing power is unleashed, memory supply buckles under the pressure...

Amidst a global memory shortage, further exacerbated by the intense demand for large memory from on-device AI, the 512GB version of the Mac Studio is already scheduled for delivery by the end of October.

Additionally, purchase limits have started.

This... Cook truly is firing off all his bullets before retiring from Apple.

Before leaving, he's given the Mac one final, massive boost!!!

2nm Debut: Apple Pumps the M6 Full of AI!

Let's start with the M6.

The reason for the emphasis is that the M6 truly ushers Apple's chips into the 2nm era!! (cheers.jpg)

The M6 is Apple's first chip manufactured using a 2nm process, expanding its CPU from the previous generation's 10 cores to 12 cores.

This includes 2 Super Cores, 4 performance cores, and 6 efficiency cores. The GPU also reaches 12 cores, and the unified memory bandwidth is increased to 170GB/s.

However, more noticeable changes are mainly reflected in the much-discussed AI capabilities

Each GPU core in the M6 directly incorporates a Neural Accelerator, paired with a Dual 16-core Neural Engine.

In other words, the GPU, originally responsible for graphics and general parallel computing, now further grows an add-on specifically for AI computations!!

Looking at these specs alone might not give much of a feel, but the data becomes more intuitive when applied to the Mac mini—

Compared to the M4 Mac mini, the AI performance of the M6 version sees up to a 4x increase, the GPU up to 2x faster, and the CPU up to 40% faster.

When running LLMs with LM Studio, prompt processing performance reaches up to 4.8 times that of the M4, and even up to 13.5 times faster compared to the original M1 Mac mini.

This time, the unified memory bandwidth also reaches 170GB/s, about 10% higher than the M5.

Although the base model still starts at 16GB, configurable up to 32GB, for running medium-to-small-sized models locally, persistent Agents, and AI workflows, the throughput capability clearly takes another step up.

This also shows that Apple's thinking about the Mac mini has somewhat changed...

Apple directly starts talking about agentic AI workflows and always-on agentic computing in its press release.

Johny Srouji even specifically mentions that the Mac mini can serve as a home computer, professional workstation, and also directly as an always-on agentic device.

The small box that used to cost around $600, quietly tucked behind a monitor for office work, is now being pushed by Apple towards being a desktop Agent mini-server.

Models can run locally, Agents can stay active, and code, files, and tasks can all remain on the machine.

It's small, but the work it handles is increasingly resembling a server.

Of course, this little box's "status" has also undergone a leap...

The 2024 M4 Mac mini started at just 4499 yuan in China, while the M6 version's starting price now directly jumps to 6999 yuan.

In two years, it's a full 2500 yuan more expensive, an increase of over 55%.

After AI capability takes off, the "small and beautiful" price of the Mac mini is starting to become not so small. (doge)

Apple's First Quad-Die: The M5 Ultra Aggressively Piles on AI Compute in a Desktop

Now onto the real heavy hitter—the M5 Ultra.

If the M6 is aggressively supplementing AI computing power for regular Macs, then with the M5 Ultra, Apple is somewhat going all-in on building a desktop AI workstation.

It's worth noting, this is also the first time Apple has implemented a quad-die architecture in the M-series SoCs.

To briefly explain, what is a quad-die architecture?—

According to the official technical description, it actually connects two sets of dual-die M5 Max chips via the new-generation UltraFusion to form a four-die system, not simply understood as four chips stacked vertically like a mille-feuille.

The performance is also quite formidable, featuring up to a 36-core CPU, 80-core GPU, 32-core Neural Engine, 512GB of unified memory, and 1.2TB/s memory bandwidth.

Peak AI performance reaches up to 4.3 times that of the M3 Ultra and 9.8 times that of the M1 Ultra; large model prompt processing in LM Studio is up to 4 times faster compared to the M3 Ultra.

There's also something very Apple and very suitable for on-device AI—Unified Memory.

On traditional PCs, CPU memory and GPU VRAM operate separately; trying to stuff several hundred GBs of VRAM into a GPU is both costly and engineeringly challenging...

But! The Mac Studio's 512GB of unified memory can be directly shared among computing units like the CPU and GPU.

This means as long as the model weights fit, many large models that previously required servers or multiple high-end GPUs can now reside directly within a single desktop machine.

And if one machine isn't enough, you can even chain more...

In a WWDC demo, Apple even gave an example using massive models: when a single model's weights don't fit into the 512GB memory of one Mac Studio, they split the weights across multiple Mac Studios to run together.

In a four-node setup, distributed AI inference performance can reach up to 3 times that of a single machine.

So by this point, the direction of this generation's Mac Studio is quite clear—

Apple is piecing together Apple Silicon + massive unified memory + MLX + Thunderbolt RDMA into a complete on-device AI computing stack.

512GB per machine, if it's not enough, just add more machines.

This is good; while others are still selling you an AI computer, Apple is starting to sell you a desktop AI server room??

AI Made Macs More Expensive, and is Also About to Make High-Memory Macs Sell Out

Problems follow.

Such powerful unified memory happens to collide with one of the hottest commodities in the AI industry in recent years: memory~

The Wall Street Journal notes that previous generations of Mac mini and Mac Studio have quietly become hot items in AI developer circles.

Especially with the rise of demand for on-device large models, Coding Agents, and tools like OpenClaw, the advantage of high-memory Macs was suddenly rediscovered—

After buying the machine, models can be run repeatedly without counting tokens for each call.

AI companies are scrambling for memory, data centers are scrambling for memory, PC manufacturers are scrambling for memory.

Now even people running Agents on Macs are joining the scramble for memory, causing a direct collision of supply and demand...

Apple's previous generation's highest memory configuration even temporarily disappeared from the official website, and while the new Mac Studio reopens the 512GB unified memory option, Apple has already warned—

This version will have to wait until the end of October, while other Mac Studios officially go on sale on September 22.

Naturally, the price isn't holding back either. For the previous generation Mac Studio, the M4 Max China model started at 16499 yuan, and the M3 Ultra started at 32999 yuan.

This generation directly becomes M5 Max starting at 19999 yuan, and M5 Ultra starting directly at 46999 yuan.

That means the Ultra entry-level model is 14000 yuan more expensive in one generation...

The price increase truly took off along with the AI computing power...

Interestingly, the person playing this card is also preparing to leave the position of Apple CEO.

On August 23, Apple just held a farewell party for Cook.

On September 1, John Ternus officially takes over.

So looking at the timeline, these two chips indeed carry a bit of the flavor of Old Man Cook's "final dance"—

In the final days of the Cook era, Apple stepped into 2nm with one hand, and with the other turned the Mac into a 512GB desktop AI prodigy.

What could be a better retirement gift for oneself?

Of course, the foldable iPhone to be released next month, while announced by the new CEO, theoretically the credit still belongs to Cook.

Steve Jobs created the iPhone, then Cook came along and... bent... folded it...

Reference Links:

[1]https://www.apple.com/newsroom/2026/08/apple-introduces-m6-and-m5-ultra-for-a-big-leap-in-performance-and-ai-compute/

[2]https://www.apple.com/newsroom/2026/08/apple-unveils-a-more-powerful-mac-mini-featuring-the-all-new-m6-and-m5-pro/

This article is from the WeChat public account "QbitAI", author: Meng Yao

İlgili Sorular

QWhat are the two main new chips announced by Apple in the article, and what is their manufacturing process?

AThe two main new chips are the M6 and the M5 Ultra. The M6 is Apple's first chip built on a 2nm process. The M5 Ultra is built on a 3nm process.

QAccording to the article, what is the most significant new architectural feature of the M5 Ultra chip?

AThe most significant new feature is its quad-die architecture (or four-die architecture), which is achieved by connecting two dual-die M5 Max chips via the next-generation UltraFusion technology.

QWhat is the maximum amount of unified memory available for the new Mac Studio with M5 Ultra, and why is this version facing supply constraints?

AThe maximum unified memory for the new Mac Studio with M5 Ultra is 512GB. This version is facing supply constraints and is delayed until late October due to a global memory shortage, exacerbated by high demand for large memory capacities to run local AI models.

QHow much does the starting price of the M6 Mac mini increase compared to the previous M4 model, and what is the percentage increase?

AThe starting price of the M6 Mac mini is 6999 RMB, which is 2500 RMB more than the 4499 RMB starting price of the M4 Mac mini. This represents an increase of over 55%.

QWhat role does the article suggest the Mac mini is evolving towards, based on its enhanced AI capabilities?

AThe article suggests the Mac mini is evolving from a small desktop computer into an 'always-on agentic device' or a desktop AI agent server, capable of locally running models and persistently handling agentic AI workflows.

İlgili Okumalar

Hubei State-Owned Assets Achieve the Largest Return in History

After years of anticipation, Yangtze Memory Holdings Co., Ltd. (YMTC) has filed for an IPO on Shanghai's STAR Market, seeking to raise 33 billion yuan—the largest offering in the board's history. This move follows the recent listing of its peer, ChangXin Memory Technologies (CXMT), which reached a market valuation exceeding 4 trillion yuan. Dubbed the "twin stars of domestic memory," both companies, founded in 2016 in Hefei and Wuhan respectively, symbolize China's push for semiconductor self-sufficiency. YMTC's journey began with its predecessor, Wuhan Xinxin, established in 2006. Backed by substantial state investment from Hubei and Wuhan, it evolved into a national memory base. The company achieved key technological breakthroughs, and now ranks as the world's third-largest and China's top NAND Flash manufacturer by sales. Its recent financials are strong, with Q1 2026 revenue of 47.04 billion yuan and net profit of 33.38 billion yuan. Post-IPO, its market value is widely expected to surpass 1 trillion yuan. The potential windfall highlights the success of long-term, patient capital from Hubei's state-owned entities. Key shareholders like Hubei Changsheng, Xintech, and government-backed funds have supported YMTC through years of development. Their collective stake could be worth hundreds of billions after the listing. This model mirrors other successes in Wuhan, such as Huagong Tech, where local state investment during a low point later yielded massive returns. The story reflects a broader national trend of regional transformation through strategic, high-tech investments. Hefei's bet on CXMT, now worth over 3.7 trillion yuan, propelled the city's A-share market cap to 4th nationally, showcasing how a major firm can reshape an entire local industry ecosystem. Similarly, Wuhan's photoelectronics cluster, now worth over 850 billion yuan, aims to become a world-class hub. The takeaway is clear: in the reshuffling of Chinese cities, patient, courageous state investment in core technologies—from memory chips to advanced manufacturing—is proving to be a decisive factor, turning long-term visions into economic reality.

marsbit13 dk önce

Hubei State-Owned Assets Achieve the Largest Return in History

marsbit13 dk önce

The Myth of AI Investment Collapses

"The AI Investment Myth Bursts: The Swift Collapse of a $45 Billion Fund The high-flying hedge fund Situational Awareness (SA), founded by 24-year-old former OpenAI researcher Leopold Aschenbrenner, neared total collapse in late July. Once a Wall Street darling, the fund saw its assets under management rocket from $1.5 billion to $45 billion in under a year, driven by a massively leveraged bet on the AI boom. Its core strategy was a 'Texas hedge'—simultaneously buying stocks seen as AI beneficiaries (like chipmakers) and shorting those deemed AI victims (like certain software firms). In reality, both sides of this trade were dependent on unbroken market confidence in AI. This strategy generated staggering returns, peaking at 439% year-to-date. However, it concealed extreme concentration, high leverage (reportedly 3-to-1), and liquidity risks from illiquid private holdings like Anthropic. When semiconductor stocks corrected sharply in late July, SA's long positions plummeted. Simultaneously, its short bets failed as 'AI victim' stocks rose, causing losses on both sides. The fund faced immediate, massive margin calls. With minutes to spare before a forced liquidation by its prime brokers, SA sold its entire public market portfolio at a discount to Citadel on July 30, narrowly avoiding a market-wide cascade. The fund's value crashed from $45 billion to roughly $10 billion (excluding its remaining Anthropic stake). The episode exposes the systemic risks embedded in the frenzied, highly leveraged chase for AI returns. It serves as a stark reminder of the old Wall Street adage: markets can stay irrational longer than investors can stay solvent. The crisis shifts focus from Aschenbrenner's AI predictions to whether capital markets will continue ignoring such dangerous concentration and leverage in pursuit of the next 'sure thing' narrative."

marsbit14 dk önce

The Myth of AI Investment Collapses

marsbit14 dk önce

İşlemler

Spot
活动图片