A Decade's Bet on Cerebras: How the 'Wafer-Scale AI Chip' Reached NASDAQ

marsbitPublished on 2026-05-15Last updated on 2026-05-15

Abstract

"Cerebras, a pioneering AI chip company, successfully debuted on NASDAQ (CBRS) on May 14, 2026, with its stock price surging approximately 68% on the first day. This marks a significant milestone following a decade-long journey, as recounted by early investor Steve Vassallo. The story begins not in 2016, but with the deep, 19-year relationship between Vassallo and founder Andrew Feldman, which started with Feldman’s previous company, SeaMicro (acquired by AMD in 2012). In 2016, Feldman and a core team of chip and system experts sought to challenge the emerging consensus. At a time when AI’s practical utility was still debated and GPUs were becoming the default hardware, they envisioned a fundamentally new computer architecture purpose-built for AI workloads. They identified memory bandwidth, not raw compute power, as the critical bottleneck for neural networks. Defying industry inertia, Cerebras pursued a radical, wafer-scale chip design—58 times larger than the biggest existing chips. This meant confronting and solving a cascade of unprecedented engineering challenges: power delivery, thermal management, and maintaining electrical continuity across tens of thousands of connections. It required reinventing nearly every aspect of modern computing—semiconductors, systems, data structures, software, and algorithms. The path was fraught with setbacks, including a prototype that caught fire on its first power-up. Progress was marked by intense, iterative problem-solving, with t...

Editor's Note: On May 14th, Cerebras officially listed on the NASDAQ under the ticker symbol CBRS. Its closing price on the first day rose approximately 68% above the issue price, making it one of the most notable AI hardware IPOs since 2026.

This article is written by Steve Vassallo, an early investor in Cerebras, who recounts his nearly nineteen-year partnership with Andrew Feldman, spanning from SeaMicro to Cerebras. On the surface, the article tells the venture capital story from term sheet to IPO. In essence, it chronicles how a frontier hardware company bet on the fundamental reconstruction of AI computing architecture during a period when consensus was skeptical: From wafer-scale chips and memory bandwidth bottlenecks to a series of engineering challenges in power supply, heat dissipation, and electrical continuity, what Cerebras faced was not a single-point technological challenge, but the re-invention of an entire modern computing system.

The most noteworthy aspect is not that Cerebras ultimately created a wafer-scale chip 58 times larger than traditional chips, but that from the outset, this company chose a direction contrary to industry inertia: When GPUs became the default answer for AI training, it attempted to redefine "what a computer designed for AI truly is." Behind this lies not only technical judgment but also the patience of capital, and, crucially, the long-term, non-transactional trust relationship between investors and the founding team.

For today's AI hardware competition, the significance of Cerebras lies in reminding the market that the compute revolution isn't just about stacking more GPUs; it may also come from re-imagining the computing architecture itself.

The following is the original text:

Friday, April 1st, 2016. I sent Andrew Feldman an email, telling him I would climb over the fence in his backyard and hand-deliver our term sheet for investing in Cerebras to him.

It was April Fools' Day, but I wasn't joking.

Strictly speaking, this wasn't standard operating procedure for a venture capital firm. But by then, I had known Andrew for nine years and had been discussing his next company with him for nearly two years. I couldn't afford to miss this deal over some sentence in the term sheet that was still being revised on a Saturday afternoon.

I first met Andrew in October 2007. At that time, he and Gary Lauterbach had just founded SeaMicro. I didn't invest in that round, but we really clicked, especially admiring their first-principles approach to problem-solving. I've been following them ever since.

Truly valuable relationships need time to mature. The same is true for truly valuable companies. Today, viewed from the outside, Cerebras is a ten-year-old company about to go public. But in my view, this is the culmination of a nineteen-year relationship, finally reaching the bell-ringing moment.

Deep Relationships, and Unreasonable Ambition

When AMD acquired SeaMicro in 2012, I had a hunch: Andrew wouldn't stay long in a big corporation. He possesses a strong unwillingness to lose and a rebellious heart. By early 2014, he was already looking for opportunities to leave, and we began meeting frequently to discuss what could be next.

At that time, two things were far from consensus: First, that AI would actually become useful; second, that GPUs were not the optimal computing architecture for AI.

Regarding the first question, many smart people I knew also disagreed. After AlexNet emerged in 2012, some corners of the research community had already begun achieving near-magical results with convolutional neural networks. But in the broader software industry, AI was still somewhere between a marketing buzzword and a research project.

The second question, the hardware question, had hardly been seriously raised. GPUs had become the default choice for neural network training, mainly because researchers accidentally discovered they were "less bad" compared to CPUs. Building a new computing system specifically for AI workloads meant challenging the mainstream architecture then being used by researchers worldwide.

But Andrew, Gary, and their co-founders Sean, Michael, and JP saw a different path. They each brought decades of experience in chips and systems: Gary's background stemmed from pioneering work on dataflow and out-of-order execution in the 1980s; Sean focused on advanced server architecture; Michael handled software and compilers; JP was deeply versed in hardware engineering. They were an exceptionally rare group: individually outstanding; collectively, their capabilities multiplied. They could imagine an entirely new kind of computer.

They believed that if AI truly unlocked its potential, the resulting market size would far exceed the sum of all existing computing paradigms.

They also saw the essence of the GPU: It was originally a chip designed for graphics processing, just temporarily promoted as an AI training tool on a new battlefield. It was indeed better at parallel processing than CPUs, but if one designed from scratch for AI workloads, no one would create an architecture like the GPU. What truly limited neural network capabilities was not raw compute power, but memory bandwidth. This meant the chip they aimed to create would not primarily optimize matrix multiplication in isolated cores, but rather how data flows efficiently throughout the entire computational structure.

Internally, investing in Cerebras was far from a consensus decision. Several of my partners had seen the previous round of semiconductor investments resulting in mostly losses, and they were very candid about their concerns. But ultimately, we agreed as a team. That weekend in April 2016, we clearly told Andrew: We wanted to be the first to give him a term sheet.

A few weeks later, Andrew, Gary, Sean, Michael, and JP moved into our EIR office space on the second floor at 250 Middlefield. I still have the floor plan the office manager drew back then. On that map, Cerebras sat next to a founder from Foundation, just a few doors away from Bhavin Shah, who later founded Moveworks. It was a good floor for startup growth.

Knowing Which Rules Can Be Bent, Which Must Be Broken

Before Cerebras, the largest chip in computing history was roughly 840 square millimeters, about the size of a postage stamp. The chip Cerebras created measures 46,000 square millimeters, 58 times larger than its predecessor.

Choosing a wafer-scale chip also meant choosing all the downstream design challenges that came with it. In the nearly 80-year history of computing, no one had truly accomplished this before. It also meant that no one had systematically solved these problems: How to power such a massive chip? How to cool it? How to maintain electrical continuity across tens of thousands of connection points?

To achieve wafer-scale computing, Cerebras essentially had to simultaneously reinvent nearly every facet of modern computing: semiconductors, systems, data structures, software, and algorithms. Each direction alone could be a startup. Andrew and his team chose to tackle the most difficult technical problems first. Through their intense, almost tireless efforts, these problems were tackled one by one.

Every six to eight weeks, we'd have a board meeting. They would walk us through what they had tried since the last meeting: a new variant of system design, a new power delivery scheme, or a thermal management adjustment. By repeatedly confronting systemic challenges head-on from every angle, they developed a hard-won clarity in articulation. They would explain where they thought things went wrong and what they planned to try next.

We would ask questions, then dive deep with the team, mobilizing the people, resources, and connections needed to help them find new approaches. Six to eight weeks later, when we met again, the story would repeat with another technical frontier: another boundary that needed exploring. Each solution would reveal the next problem that had to be solved.

Their first prototype wafer literally smoked the first time they powered it on. The team called it a "thermal event"—what you call a fire when you don't want to scare the board or the landlord.

I had been calculating power consumption per square millimeter, partly out of curiosity, partly because the numbers seemed too high to be true. So, we brought in engineers from Exponent, a failure analysis firm whose former company name was, aptly, Failure Analysis. They confirmed that the power numbers were indeed as audacious as they appeared and helped us think through options that didn't challenge the second law of thermodynamics. After all, that was one law Andrew was smart enough not to argue with.

The discipline of an engineer lies in knowing which rules can be broken, which can be bent, and which must be respected. Andrew and his team had a practiced intuition for that distinction. They knew when they were challenging convention—which they intended to do—and when they were challenging physics—which they did not.

When you're building frontier technology, failure is inevitable. The only way through it is discipline, persistence, and most importantly, trust: trust in the mission, trust in each other, and trust in the idea that when the first prototype self-destructs, you'll all be back in the lab the next morning for the next iteration.

There's no transactional version of this work. There's only the long-term version: staying in the room, through the incomplete solutions and patient explanations, so that when it finally works, you are there to see it.

That moment arrived in August 2019. Andrew, Sean, and their team stood in the lab, watching a new computer they had designed from scratch run for the first time. To an outsider, it superficially didn't seem to be doing anything interesting. According to Andrew, it was probably about as exciting as watching paint dry. The difference this time was: no bucket of "paint" like this had ever dried before. They stood there together for 30 minutes, then went back to work.

Who You Build With, Matters

Some people choose problems based on what they know they can solve. Andrew's criteria for choosing problems is what he believes is worth solving. Incremental iteration doesn't excite him; he wants 1000x leaps. From day one, he wanted to build Cerebras into a generational, one-of-a-kind company.

Part of that drive comes from his personality. Andrew describes it as a computer architect's "disease"—being haunted by an idea for decades. But to me, it's more broadly a founder's "disease." He looks at a problem and first asks himself: Can I make something that causes a step-function improvement? Then he asks: If I succeed, will anyone care? If the answer to both is yes, he will commit the next decade of his life to it.

Another part of that drive comes from his upbringing. Andrew grew up surrounded by geniuses as naturally as most kids grow up watching TV. His father was a pioneering evolutionary biology professor who played rotating doubles tennis every Sunday with six other people. Three of those six later won Nobel Prizes, and one won a Fields Medal.

According to Andrew, these giants would patiently explain their work in physics, mathematics, and molecular biology to him in language a child could understand. He formed a deep impression of what true intelligence looks like and also understood, as his mother said, that being smart doesn't mean you have to be a jerk.

I've come to realize this is one of Andrew's core traits, as important as his rebellious ambition and his almost phototropic instinct for truly worthy problems. He deeply believes that the most exceptional people he's encountered are also often extraordinarily kind.

This belief shaped how his team came together to accomplish incredibly hard things. The first 30 people Cerebras hired had all worked with him before; some had been with him since 1996. Today, Cerebras has about 700 employees, and roughly 100 of them have followed him across multiple companies.

The important thing is, kindness and competitiveness are not mutually exclusive. Andrew has an intense desire to win. He likes to say he's a professional version of David, fighting Goliath. Goliath is slow-moving and always guarding against frontal attacks, which leaves room for every other move. David's advantage lies in showing up in ways and places Goliath cannot.

At SeaMicro, Andrew's largest channel partner in Japan was NetOne. NetOne's primary supplier was Cisco, which would entertain partners with private jets and yachts worth more than most houses in Palo Alto. Andrew's budget was far more modest, so he invited NetOne's CEO to his backyard for a barbecue. Later, the CEO told him he had done business with Cisco for decades but had never been invited to anyone's home. That seemingly small, very human gesture—something a Goliath would never think to do—cemented their relationship.

From the First Term Sheet to IPO

This morning, Andrew rang the opening bell at NASDAQ. I stood next to him. It's been ten years and 2600 miles since it all began in our 250 Middlefield office.

Today, there are still rare founders doing what Andrew did: sketching on whiteboards at 3 a.m., wrestling with technical problems not yet solved. They also harbor a strong unwillingness to lose and a rebellious heart. They are trying to find a partner who is truly willing to work side-by-side: willing to dive in and help solve the problem when the first prototype won't power on; and who will stay until it finally runs.

These are precisely the founders I want to back: those who choose problems worth solving, imagine a solution 1000x better than the status quo, and persistently hone and persevere through the inevitable challenges along the way.

For founders like Andrew, Gary, Sean, Michael, and JP, I'm willing to climb over a backyard fence on a Saturday afternoon to hand-deliver a term sheet.

Trending Cryptos

Related Questions

QWhat was the core technical challenge and architectural vision that Cerebras pursued, as opposed to the industry consensus?

ACerebras challenged the industry consensus that GPUs were the optimal architecture for AI training. While GPUs became the default due to their superior parallel processing compared to CPUs, the Cerebras team believed they were not designed for AI from first principles. Their core architectural vision was to design a computer specifically for AI workloads by fundamentally addressing the memory bandwidth bottleneck, not just raw compute power. This led them to invent the wafer-scale chip, a system 58 times larger than the largest previous chips, which required re-inventing nearly every aspect of modern computing—semiconductors, power delivery, cooling, and software—to enable efficient data flow across the entire compute structure.

QHow does the article characterize the relationship between the investor (Steve Vassallo) and the founder (Andrew Feldman), and why was it crucial for Cerebras's journey?

AThe article characterizes the relationship as a deep, long-term, non-transactional partnership built on trust over nearly two decades. This relationship was crucial because building Cerebras involved tackling a series of unprecedented engineering failures (like the first prototype catching fire) and systemic challenges over many years. The investor's patience, willingness to engage deeply with technical setbacks during bi-monthly board meetings, and commitment to providing resources and relationships allowed the founder and team to persist through iterative failures without pressure for short-term results. This trust-based support system was essential for navigating the 'inevitable' failures of frontier technology development.

QAccording to the article, what are the key personality traits and background influences that shaped Andrew Feldman as a founder?

AAndrew Feldman is described as having a strong不服输 (refusal to accept defeat) and a rebellious heart. He is driven by a desire for 1000x leaps rather than incremental improvements and is drawn to solving problems he believes are truly worth solving. Key traits include: a 'founder's disease' of being obsessed with a transformative idea; a competitive spirit, seeing himself as a 'professional David' against Goliaths; and a core belief that the most brilliant people are also kind, a value instilled by his mother. His background growing up surrounded by intellectual giants (including future Nobel laureates) who were patient and kind gave him a model of excellence coupled with decency, which influenced how he built and led his teams with loyalty and humanity.

QWhat does the Cerebras story suggest about the nature of innovation in AI hardware beyond simply using more GPUs?

AThe Cerebras story suggests that true innovation in AI hardware requires a fundamental re-imagination of computing architecture itself, not just scaling existing solutions like GPUs. It demonstrates that a compute revolution can come from addressing foundational bottlenecks like memory bandwidth and designing a system from the ground up for a specific workload (AI), rather than adapting a tool designed for another purpose (graphics). This path involves tackling a holistic set of interdependent engineering challenges—power, cooling, electrical continuity, software—that constitute 're-inventing the modern computing system.' It underscores that such innovation demands long-term capital patience, technical judgment, and a willingness to pursue a direction contrary to industry inertia.

QWhat symbolic and practical significance did the act of delivering the term sheet over a backyard fence hold, as described in the article?

AThe act of delivering the term sheet by climbing over Andrew Feldman's backyard fence on a Saturday held both symbolic and practical significance. Symbolically, it represented the investor's exceptional commitment, personal dedication, and willingness to go beyond standard venture capital protocols for a founder and a vision he deeply believed in. Practically, it underscored the urgency and importance of securing the deal—the investor did not want to miss the opportunity due to last-minute term sheet edits. This gesture foreshadowed the long-term, hands-on, and trust-based partnership that would be essential for navigating Cerebras's decade-long journey of overcoming seemingly impossible technical hurdles.

Related Reads

From Gold to Bitcoin: Fixed Supply + Institutional Frenzy, Might It Repeat the 'Explosive' Price Trend?

"From Gold to Bitcoin: Fixed Supply and Institutional Frenzy May Lead to 'Explosive' Price Rally Analysts suggest Bitcoin's price action could mirror gold's over the past two decades, following the launch of spot Bitcoin ETFs. Gold ETFs, introduced in 2004, drove gold's price surge to a current market cap near $28 trillion. Both gold and Bitcoin are non-yielding stores of value, with prices driven purely by investor sentiment rather than cash flows or credit. Gold ETFs experienced dramatic cycles: explosive growth, painful drawdowns, and slow recoveries, with each cycle reaching higher peaks. Bitcoin ETFs, approved in early 2024, saw rapid institutional adoption but are now facing similar volatility. Recent warnings highlight the risk of significant ETF outflows disrupting the current rebound. BlackRock's IBIT, a leading Bitcoin ETF, has sold nearly 100,000 BTC to meet redemptions while still holding over 733,000. The core parallel is fixed supply: when demand surges, prices explode, but demand is often volatile and wave-like, not steady. Institutional interest, through ETFs and corporate adoption, remains a key support pillar, helping to cushion sell-offs. If Bitcoin captures even a fraction of gold's role as a store of value, its upside potential is immense, though the path will be marked by high volatility. For investors, focusing on long-term trends and managing risk is crucial as this 'price explosion' narrative unfolds."

Foresight News14m ago

From Gold to Bitcoin: Fixed Supply + Institutional Frenzy, Might It Repeat the 'Explosive' Price Trend?

Foresight News14m ago

Why Is AI Agent Shopping Hard to Popularize?

The article argues that the popular narrative of "AI agent shopping" – equipping AI with a wallet to autonomously handle purchases – is fundamentally flawed and oversimplifies the complexity of shopping. It deconstructs shopping into two core actions: **information retrieval** (standardized, easily automated) and **value judgment** (deeply subjective and human-centric). The narrative mistakenly assumes AI can fully handle both. Value judgment itself has two layers: **evaluation** (assessing options against criteria) and **demand definition** (setting the criteria, weights, and values). The latter is inherently human and dynamic, as preferences are not fixed but constructed during the decision-making process ("constructive preferences"). The real dividing line for automation is not product standardization, but whether the **act of choosing** itself holds experiential value. For mundane purchases (e.g., printer paper), full AI delegation works. For experiential goods (e.g., wine, furniture), the joy of selection is core to consumption, so AI should act as an assistant that narrows options, leaving the final choice to humans. The "AI wallet" concept confuses three separate elements: decision-making, execution, and fund custody. Current payment industry solutions (e.g., from Stripe, Mastercard, Google, Visa) show that limited, scoped payment authorization tokens are sufficient for most consumer scenarios, not full fund custody. The true use case for autonomous AI wallets is in **B2B procurement** and **machine-to-machine (M2M) settlements** for standardized, high-frequency, low-value transactions. The real bottlenecks for AI shopping are not payment technology, but **1) the lack of trusted data sources** (e.g., fake reviews, counterfeit goods) and **2) the impossibility of automating human demand definition**. The conclusion is that the focus should be on safely automating the assessment and filtering process while reserving for humans the rights to define their criteria and enjoy the final act of choice. For experiential goods, the platform's competitive advantage shifts to providing a superior selection experience.

Foresight News1h ago

Why Is AI Agent Shopping Hard to Popularize?

Foresight News1h ago

After Nine Months of Shorting, a Full Turn to Long: Renowned Trader Opens Bitcoin Positions Around 64K, Crypto Market Long-Short Divergence Intensifies

After nine months of being short, prominent crypto trader Doctor Profit has closed all his bearish positions and started buying Bitcoin near $64,000, signaling a complete bullish reversal. He argues that structural market changes—such as impending U.S. regulation (CLARITY Act) and institutional adoption via securities tokenization—are rewriting the traditional four-year cycle script, potentially bringing the market bottom forward from the widely expected September/October timeframe. This view finds some technical support from on-chain analyst gumsays, who notes a bullish divergence on Bitcoin's weekly chart has persisted for 147 days, nearing the 161-day duration seen before the 2022 cycle low. However, cycle researcher Jake Pahor presents a counter-argument based on historical data. Analyzing patterns since 2014, he identifies three common features of past bear market bottoms: a ~12-month duration from peak to trough, a sustained period of extreme fear (with a proprietary risk score below 20), and the price falling below Bitcoin's realized price (~$53,000 currently). The current cycle, only nine months from its October 2025 peak, meets none of these conditions. The debate highlights a market torn between "front-running" a potential early bottom driven by new fundamentals and waiting for confirmation through traditional on-chain and sentiment metrics. While Doctor Profit opts for aggressive buying, Pahor maintains a disciplined, tiered accumulation strategy, continuing weekly buys at current risk levels but reserving larger orders for if more extreme fear emerges.

marsbit1h ago

After Nine Months of Shorting, a Full Turn to Long: Renowned Trader Opens Bitcoin Positions Around 64K, Crypto Market Long-Short Divergence Intensifies

marsbit1h ago

Senior Trader's Confession: How to Trade Market's False Expectations?

Veteran trader's case study: trading the market's "wrong expectations". This trade centered on a textbook "expectation error" after a weak CPI report. While the market initially priced in broad monetary easing (sending Nasdaq to 30,060), the crucial 30-year real yield hit a 20-year high. This signaled a fractured transmission mechanism: short-term rates eased, but long-term funding costs (vital for tech valuations) refused to fall. The trader executed five short positions on the Nasdaq (NQ) as it fell from 30,060 to 28,768. The core methodology: don't just trade the data, but analyze the market's implied causal chain and identify where it breaks. In this case, the chain was: Weak CPI → Policy Easing → Lower Long-Term Funding Costs → NQ Valuation Expansion. The break occurred between policy easing and long-term rates. The "veto variable" – long-term real yields – refused to confirm the bullish narrative. Trades were structured around "fast variables" (price) temporarily repairing while "slow variables" (funding conditions) remained broken. The article outlines a repeatable framework: 1) Map the market's implied causal chain. 2) Identify the veto variable. 3) Observe if it rejects the narrative. 4) Enter when price still follows the old script. 5) Choose the cleanest asset expression (e.g., short NQ, not broad S&P). 6) Define both invalidation and fulfillment exit conditions. The key insight: Alpha often comes not from an information edge, but from a "reaction function edge" – recognizing when the market is applying an outdated causal logic to new data. The critical question: What causal chain is the market's first reaction relying on, and is that chain still valid today?

marsbit2h ago

Senior Trader's Confession: How to Trade Market's False Expectations?

marsbit2h ago

Trading

Spot

Hot Articles

Discussions

Welcome to the HTX Community. Here, you can stay informed about the latest platform developments and gain access to professional market insights. Users' opinions on the price of S (S) are presented below.

活动图片