ChatGPT's New Model Bel Rumored to Have Completed Pre-training with a Staggering 10 Trillion Parameters

marsbit2026-08-31 tarihinde yayınlandı2026-08-31 tarihinde güncellendi

Özet

ChatGPT's new model "Bel," reportedly pre-trained with a staggering 10 trillion parameters, signals OpenAI's aggressive push toward AGI. This "monster" model, succeeding the "Doug" base model and targeting the post-GPT-6 era, allegedly surpasses the upcoming Astra model in coding, reasoning, and long-horizon agent tasks. It can operate autonomously for days, self-recover, and coordinate hundreds of sub-agents. The report claims OpenAI has consolidated computing power, even pausing projects like Sora, to focus on these massive models. With its proprietary "Jalapeño" chips and significant computational advantage, OpenAI reportedly believes competitors like Anthropic cannot keep pace. Internally, OpenAI is said to be achieving recursive self-improvement (RSI), using AI to optimize its own systems and drastically cut costs. Future visions include ultra-fast, low-latency AI enabling real-time "flow state" interaction, a shift to cloud-based agent clusters over local compute, and the emergence of a "Personal AGI" that passively understands and proactively assists users. While the specifics of Bel remain unconfirmed, the leaks portray an AI race where scale and compute are key, with OpenAI positioning itself to maintain a dominant lead through 2026-2027.

The tech world is losing sleep!

OpenAI's most powerful model, Astra, is rumored to possibly launch next Thursday, and its testing scope has already been expanded.

Previously, internal secrets were leaked, revealing that OpenAI has formally taken over the underlying code of its in-house chip "Jalapeño." AI-written core code runs 1.8 times faster than that of top human engineers. Signs of AI self-recursion have emerged within OpenAI!

OpenAI has another depth charge.

Rumors suggest OpenAI has successfully run the pre-training for a model codenamed "Bel" with 10 trillion parameters, directly targeting the ultimate threshold of AGI. It's even said that the 10-trillion-parameter pre-training is just the starting point for Bel. After that, Bel can learn at two speeds.

Altman has declared: We should throw another party for the next-generation model release.

What ambitions lie in the world beyond GPT-6?

Bel: The Abyssal Leviathan with 10 Trillion Parameters

This year, OpenAI seemed to hit a scaling bottleneck, even triggering an emergency "Red Code" alert.

To concentrate computing power, OpenAI even cut products like Sora and the AI browser Altas.

The next-generation AI model Astra achieved several mathematical breakthroughs, stunning the world, but its release has been delayed.

In recent weeks, OpenAI saw personnel upheaval: Chief Revenue Officer Dennis Dresser, Chief Operating Officer Brad Lightcap, and Fergie Simo, former deputy to CEO Sam Altman, among other executives, left one after another.

Just as outsiders speculated whether OpenAI had "run out of talent," several hardcore tech insiders revealed that OpenAI just completed a super-large-scale pre-training, codenamed "Bel."

How terrifying is this "Bel"?

Let's look at a few key terms:

1. Breaking the 10T (10 Trillion) Parameter Barrier

If the trillion-parameter GPT-4 gave AI common sense and logic close to that of a human undergraduate, what kind of "emergent abilities" will the 10-trillion-parameter Bel exhibit in complex reasoning, long-text association, and cross-disciplinary multimodal understanding?

This is a qualitative change.

It's equivalent to packing together the brain capacity of all the world's top experts and multiplying it by an exponential amplifier. At this parameter scale, AI's understanding of world models will reach an unprecedented depth.

2. The Successor to "Doug," the Ultimate Foundation Model Post-GPT-6

Insiders revealed that prior to this, OpenAI had already completed pre-training for a model codenamed "Doug."

Doug was positioned as the base model for the Astra project and the rumored GPT-6 (to be followed by extremely intensive reinforcement learning alignment).

As its successor, Bel is the next-generation foundation model that goes a step beyond Doug, belonging to the "post-GPT-6 era."

Bel directly aims at the tech world's holy grail—AGI (Artificial General Intelligence).

Reportedly, the Bel model surpasses Astra in coding, reasoning, and long-duration agent tasks.

Sources claim the model can operate continuously and efficiently for days without intervention, self-recover, and coordinate hundreds of parallel sub-agents.

3. The Claude Fable Killer (The Fable Killer)

@ChrisGPT stated bluntly in his tweet:

Bel is OpenAI's "monster" model, designed to be the Fable killer.

It should arrive by the end of this year or within a few months of Astra's release.

He even received this codename six days ago, corroborating the source's reliability.

Theoretically, Bel may have already surpassed GPT-6, even approaching OpenAI's defined AGI threshold.

In the official release of GPT-5.6, they introduced the "RSI index," which integrates achievements in research debugging, kernel and training recipe optimization, machine learning experiments, and model self-improvement. Ultimately, the sol model improved by 16.2 points over GPT-5.5.

Subsequently, sol designed hundreds of architecture experiments for its smaller draft models and initiated training. Human intervention only occurred in cases of hardware failure or training instability. Ultimately, token generation efficiency improved by over 15%.

Bel becomes a truly continuously evolving entity.

Fast weight layers absorb lessons learned during its operation from verified proofs, code tests, experiments, and tool trajectories. Slower cycles consolidate improvements that survive evaluation into persistent weights and training recipes.

It learns rapidly in fast memory, solidifies validated improvements into slow weights, continuously optimizes the operational mechanisms for the next learning cycle, and distills the final outcomes into smaller, practical models for everyone to use.

GPT-7 might just be a safe snapshot of Bel's state in a particular week.

On Reddit, this message from Leo has sparked extensive discussion, given Leo's reputation for reliable leaks.

Some speculate that internal models might be 4.5-6 months ahead of external models.

OpenAI Declares: "Anthropic Can't Keep Up Anymore"

If Bel is a dimensional reduction strike in technology, then computing power is OpenAI's secret weapon for soaring internal morale.

According to cross-analysis by multiple trackers, OpenAI judges that maintaining the lead is unquestionable for the second half of 2026 through 2027.

Why such certainty? Because of computing power.

The leak mentions that OpenAI's internal assessment believes its greatest rival, Anthropic, is short on computing power and struggles to compete with OpenAI's next-generation AI.

In the arms race of large models, computing power is ammunition. When model parameters soar to the 10-trillion level, a single training run incurs colossal costs.

While Anthropic has extremely high achievements in model architecture and alignment technology, it clearly struggles in the face of absolute "brute force aesthetics."

Facing Astra's impending public debut, constrained by computing bottlenecks, Anthropic will likely find it difficult to mount a strong response within this year.

While other companies are still scrambling to assemble 100,000 H100/B200 chips, OpenAI has already smashed out "Doug" and "Bel" with brute force aesthetics.

And with the release of OpenAI's in-house chip Jalapeño, the moat hasn't been filled; it has been widened.

OpenAI Codex Lead Reveals "Endgame Vision"

Ten trillion parameters might seem distant. But within OpenAI, these foundational brute-force breakthroughs all point towards the same AI endgame.

Recently, on the popular tech podcast with Matthew Berman, OpenAI executive Tibo unreservedly teased OpenAI's future roadmap at the application layer.

He even boldly stated:

The incredibly powerful Codex model of today will seem like a primitive artifact in just 2 to 3 months.

Combined with earlier leaks, Tibo points to four disruptive upheavals.

Recursive Self-Improvement (RSI) happens daily inside OpenAI.

Tibo confirmed that the "internal singularity" not only exists but has already achieved a commercial closed loop.

OpenAI has long been using its strongest models to optimize its own inference stack, CUDA kernels, and even infrastructure.

He gave an example: OpenAI internally used the Sol model to optimize the Luna model, directly slashing operating costs by 80%!

"Ultra Fast" Will Reshape Human Flow State.

Tibo revealed that internally, the current Ultra Fast mode has achieved up to 14x acceleration. He predicts that within just 1 to 2 years, such ultra-low latency will become the industry default standard.

When AI's response speed approaches or even surpasses human thinking speed, interaction will become real-time. Workflows will completely return to the "Flow State," and human cognitive load will decrease exponentially.

Computing Paradigm Shift: Your PC is About to Become "Scrap Metal."

With the arrival of the next-generation models (like Astra), Tibo clearly stated: The future mainstream of AI will absolutely not be local execution but will completely shift to large-scale Agent clusters in the cloud.

The Final Blow: ChatGPT and Codex Merge, Transforming into "Personal AGI."

In the future, "the mechanism will completely disappear." You will no longer need to manually write complex prompts, maintain skill files, manage memory, or manually schedule sub-agents.

Personal AGI will continuously and passively understand your goals, daily habits, and team dynamics, and proactively offer assistance.

Even more remarkable is its dynamic UI adaptation. The underlying technology is the same multimodal, Voice-first tech, but the interface will automatically morph like water based on your identity.

Conclusion: The Singularity is Here, Invisibly Present

Regardless of how exaggerated the rumors about "Bel" are, or whether Astra really shreds human expert code so smoothly, this large-scale leak has sent an undeniable signal to the world:

The development of artificial intelligence has not stagnated; it is merely gathering the momentum for a storm powerful enough to flip the table.

The second half of 2026 is just the beginning of the spectacle.

References:

https://x.com/ChrisGPT/status/2092334431142850782

https://x.com/fanofaliens/status/2091805645221789771

https://x.com/notjazii/status/2092336615473701361

https://www.reddit.com/r/singularity/comments/1vy99vk/according_to_leo_openai_just_finished_its_next/, https://x.com/imjustnewatai/status/2092482501524516888?s=20

This article is from the WeChat public account "New Zhiyuan" (ID: AI_era), author: David

İlgili Sorular

QWhat are the key features of OpenAI's reported new model 'Bel'?

AAccording to the article, OpenAI's 'Bel' model is reported to have key features including: 1. A massive scale of over 10 trillion parameters. 2. It is positioned as the successor to 'Doug', aiming to be the foundational model for the post-GPT-6 era and targeting AGI (Artificial General Intelligence). 3. It is described as a 'Claude Fable killer' and is expected to be released around the end of the year or a few months after Astra. 4. It allegedly exhibits capabilities like self-recovery, coordinating hundreds of sub-agents, and learning continuously at two speeds.

QWhat is the 'Doug' model mentioned in the article?

AThe article states that 'Doug' is a pre-trained model completed by OpenAI before 'Bel'. It is positioned as the base model for the Astra project and the rumored GPT-6. It is a precursor model, with 'Bel' being its more advanced successor aimed at the post-GPT-6 era.

QAccording to the article, why is OpenAI confident about its lead over Anthropic?

AThe article cites OpenAI's internal assessment, as per sources, that its biggest rival, Anthropic, is facing a significant compute power shortage. OpenAI believes this puts Anthropic at a disadvantage in keeping up with OpenAI's next-generation AI models like 'Bel' and 'Astra', especially given the massive compute requirements for training 10-trillion-parameter models. OpenAI's access to vast computational resources and its upcoming custom chip 'Jalapeño' are seen as key advantages.

QWhat future AI application trends did OpenAI's Tibo reveal in the interview?

AIn an interview, OpenAI's Tibo outlined several future application trends: 1. Recursive Self-Improvement (RSI) is already happening internally, optimizing infrastructure and reducing costs. 2. 'Ultra Fast' interaction speeds (up to 14x faster) will become the industry standard, enhancing human 'flow state'. 3. A paradigm shift from local AI to cloud-based large-scale Agent clusters. 4. The fusion of ChatGPT and Codex into a 'Personal AGI' that proactively assists users without manual prompting or agent management.

QWhat is the significance of the '10 trillion parameters' mentioned for the Bel model?

AThe '10 trillion parameters' figure for the Bel model signifies a monumental leap in scale. The article suggests that while GPT-4's trillion parameters gave AI near-undergraduate-level reasoning, scaling to 10 trillion parameters represents a qualitative change ('qualitative leap'). It is expected to lead to unprecedented 'emergent capabilities' in complex reasoning, long-context understanding, and cross-disciplinary multimodal comprehension, potentially bringing the model closer to the threshold of AGI (Artificial General Intelligence).

İlgili Okumalar

Marvell: Can't Compare to NVIDIA, Can't Meet Expectations, Overvaluation Gets Squeezed First?

Marvell Technology (MRVL.O) reported its Q2 FY2027 earnings (ending July 2026) after market close on August 27. Key points include: The company raised its full-year revenue outlook for FY2027 to $12 billion (from $11.5B) and for FY2028 to $18 billion (from $16.5B). However, these upward revisions were only slightly above market expectations and significantly trailed NVIDIA's recent explosive guidance. The Data Center segment, accounting for 79% of revenue, grew 19% quarter-over-quarter to $2.17 billion, primarily driven by connectivity products. For FY2028, management forecasts over 60% growth for this segment, again below NVIDIA's >70% outlook. A major disappointment for investors was the lack of an upward revision to the Custom ASIC business guidance, despite Marvell's recent partnership agreement with Google. The market had anticipated potential gains from Google's TPU orders, but the maintained guidance for "over 100% growth" in FY2028 (with no specific target for FY2027) led to concerns that the Google deal may be a less favorable "framework agreement" where Marvell holds a weaker negotiating position. Adjusted gross margin was flat at 58.3%. Q3 revenue guidance is $3.15 billion, slightly above consensus. Overall, the report was largely in line with expectations, but the subsequent stock decline is attributed to growth forecasts that failed to meet heightened market expectations (particularly versus NVIDIA) and lingering uncertainty around the tangible benefits of the Google ASIC partnership. High valuation faces near-term pressure, but expectations for >50% growth in the coming years and long-term ASIC opportunity may provide support.

marsbit12 dk önce

Marvell: Can't Compare to NVIDIA, Can't Meet Expectations, Overvaluation Gets Squeezed First?

marsbit12 dk önce

US Stock Market Trend (August 31st): Kashkari's Hawkish Remarks Weigh on Chip Stocks, US-Iran Weekend Strikes Boost Oil Prices

U.S. stock markets ended lower on Friday following hawkish remarks from Federal Reserve Chair Wash at the Jackson Hole symposium, which sharply increased the probability of a September rate hike from 35% to nearly 60%. Major indexes fell: the S&P 500 dropped 0.25%, the Nasdaq declined 0.52%, and the Dow was essentially flat. This shift in interest rate expectations pressured rate-sensitive assets, leading to significant declines in chip stocks. The Philadelphia Semiconductor Index fell 3.47%, with Nvidia dropping 4.57%, erasing about half its post-earnings gains. Geopolitical tensions also escalated over the weekend as the U.S. and Iran exchanged military strikes, raising concerns over the security of oil transit through the Strait of Hormuz. This pushed oil prices up over 2% in early Asian trading on Monday, reintroducing a geopolitical risk premium. In other energy news, former President Trump announced a landmark 25-year oil deal with Venezuela, aiming to significantly increase the country's oil production. However, this long-term supply boost was overshadowed in the short term by the Middle East conflict and the dominant market focus on interest rates. The core market narrative for the coming week revolves around the interplay between re-priced hawkish rate expectations and escalating geopolitical risks. Key areas to watch include the trajectory of Treasury yields, the evolution of U.S.-Iran tensions and its impact on oil prices, and whether the sell-off in high-valuation tech and semiconductor stocks stabilizes or continues under the pressure of higher rates.

marsbit28 dk önce

US Stock Market Trend (August 31st): Kashkari's Hawkish Remarks Weigh on Chip Stocks, US-Iran Weekend Strikes Boost Oil Prices

marsbit28 dk önce

a16z: Top Talent Flows to AI Infrastructure, Infrastructure Design Will Be 'Redesigned from Scratch'

a16z Unveils "Machine Age Fund": AI Infrastructure Faces Massive Overhaul Silicon Valley VC giant a16z (Andreessen Horowitz) has launched a new "Machine Age Fund" dedicated to AI infrastructure, citing a vast and growing "supply-demand fracture." Key takeaways: * **Unlimited Demand vs. Constrained Supply:** AI demand is growing exponentially (estimated near 1000% annually for tokens), while supply chains for chips, memory, data centers, and power are booked through 2027-2028. GPU prices are rising against historical trends. * **A Resource Problem, Not Engineering:** The bottleneck is no longer software engineering but physical resources (hardware, power, cooling). Money and compute directly translate to intelligence output, removing traditional scaling limits. * **Complete Infrastructure Rebuild Needed:** Existing data centers and computing stacks, designed for a different era, are hitting physical limits. Everything needs rethinking from first principles: chip architecture, memory hierarchy, networking, power delivery (shifting to 800V DC), and cooling (moving to liquid). * **Investor & Founder Shift:** Top entrepreneurs are increasingly moving into hardware, with deals in the space rising from ~3-5% to over 20-30% of a16z's top-tier deal flow. Founders need to be "systems thinkers" who understand manufacturing and supply chains. * **Massive Economic Scale:** Training a frontier model now costs $3-5B. With inference needing to recoup ~$10B, saving 20% in efficiency ($2B) can justify developing a custom ASIC for a single model—a previously unthinkable economic dynamic. * **Long-Term Horizon:** a16z believes we are in the very early stages of a decades-long era where compute is applied to vast new domains (science, materials, biology, creative work). The firm re-frames AI as "Machine Intelligence," emphasizing the critical, foundational role of hardware in this new age.

marsbit1 saat önce

a16z: Top Talent Flows to AI Infrastructure, Infrastructure Design Will Be 'Redesigned from Scratch'

marsbit1 saat önce

Claude's Safety Mechanism Backfires Spectacularly, AI Deletes Developer's 700GB Home Directory in a Fit of Rage

Claude's safety mechanisms backfired spectacularly, leading an AI to delete a developer's entire 700GB home directory. The developer, Guillemot, asked Claude Fable 5 to write a script to create isolated sandbox folders in /tmp for AI agents and clean them up afterward, ensuring files in use wouldn't be deleted. Because the task involved risky delete operations, Claude's built-in safety system triggered an "adversarial review," automatically downgrading the model from the capable Fable 5 to the more conservative Opus 5, and then to Opus 4.8, to perform a security check. The Opus 4.8 model successfully tested the script, correctly identifying the /tmp and home directories as off-limits for deletion. However, during the cleanup phase that followed the test, it reused a variable name from the testing phase. This variable contained the path to the user's home directory. Consequently, the model executed a delete command on that exact path it had just deemed unsafe, wiping out 700GB of data. The intended /tmp directory remained untouched. The incident highlights critical flaws in Anthropic's safety downgrade system, which community developers have criticized for being overly sensitive, reducing model capability for complex tasks, and being "sticky" for an entire session once triggered. The irony is that a mechanism designed to hand off dangerous tasks to a safer, more conservative model instead delegated it to a weaker model that made a catastrophic error in variable scope and path handling.

marsbit1 saat önce

Claude's Safety Mechanism Backfires Spectacularly, AI Deletes Developer's 700GB Home Directory in a Fit of Rage

marsbit1 saat önce

İşlemler

Spot
活动图片