Token Plans Launch: The 'Traffic War' in the AI Era, Now It's the Turn of Doubaos to Compete

marsbitPublicado em 2026-05-19Última atualização em 2026-05-19

Resumo

China's major telecom operators are launching standardized "Token" service packages, marking a new phase in the AI era where model usage is becoming a commodity akin to mobile data plans. Operators like China Telecom and China Mobile are offering monthly subscription plans for individuals and enterprises, allowing access to dozens of AI models through unified platforms at set Token rates (e.g., 9.9 yuan for 10 million Tokens). This shift lowers the cost and technical barrier for users to switch between models like Doubao, Qwen, and DeepSeek. The article explains that a Token is the basic computational unit for large language models. Operators are transforming from selling voice minutes and data bandwidth to selling AI compute measured in Tokens. This model benefits developers and SMEs by providing predictable costs and easy access to multiple models without managing underlying infrastructure. As operators become aggregation platforms, competition among model providers intensifies. They must now compete not just on model capability but also on price, computational efficiency (cost per quality Token), and higher-value AI application solutions. The future may see a split where operators control the user access point, while model companies focus on core AI capabilities and specialized enterprise applications.

By Silicon Base Quadrant

When users are no longer debating whether to upgrade their monthly data plan, they may soon start debating how many Token services to purchase each month.

Tokens are about to be packaged and sold as standardized services by telecom operators, just like data, broadband, and SMS.

Recently, China's three major telecom operators have successively launched Token plan products: monthly subscription-based Token schemes for individual users, and tiered computing power packages for developers and enterprise customers. They have announced the integration of dozens to hundreds of large models into their platforms, allowing for "monthly purchase, multi-model calls, and payment via phone bill."

China Telecom has launched personal and enterprise Token plans, with a minimum monthly fee of 9.9 yuan for 10 million Tokens; local operators like Shanghai Mobile and Shanghai Telecom have introduced billing models based on quota points or general Tokens, with Shanghai Mobile offering 400,000 Tokens for 1 yuan.

As operators begin selling Token services, the cost for users to switch between large models will significantly decrease. For large model companies, this means "user stickiness" will be weakened, and only by "competing more fiercely" can they retain their market.

In the future, large model vendors like Doubao, Qwen, and DeepSeek will not only compete on "price" and "Token quality per unit of energy consumption," but also on "higher-value AI application solution capabilities."

01 What is Token Service?

To understand Token service, first understand what a Token is.

Computers cannot directly recognize text; they can only recognize 0 and 1 codes. Therefore, every word, character, punctuation, or piece of speech we input is converted into 0 and 1 codes through a specific encoding mechanism.

In the context of large models, they also first recognize numeric codes, and the number of bits in the code converted from each character varies slightly.

A Token is the smallest unit of computation for a large model to process information. User input, context memory, and model output are all calculated in Tokens. More complex model calls, longer contexts, and deeper Agent execution chains consume more Tokens.

Typically: In English, one Token is roughly equivalent to 4 letters; In Chinese, due to the higher information density of Chinese characters, one Chinese character, one punctuation mark, or one phrase often corresponds to 1 to 2 Tokens.

Since large models think and output Token by Token, the industry sells and settles the cost and usage quota of large models in the form of "Per Million Tokens" or "quota points."

Currently, large model companies implement tiered pricing for Tokens. Ordinary users using general modes of models like Doubao or Qwen are free; for enterprise-level heavy usage, one can purchase different tiers of API monthly packages or metered services.

Starting last year, operators opened large model "computing power supermarkets." Model vendors are the "tenant merchants," and operators charge "platform fees + computing power fees + channel fees." What users buy is not the "operator's model," but rather: on the telecom platform, using telecom computing power, to call any large model, billed per Token.

In July 2025, China Mobile launched the model service platform MoMA (Mobile Model Access); in April, China Telecom launched the Xingchen TokenHub operation service platform; in May, "Unicom Xingluo" Token service platform was released. These platforms have integrated mainstream large models from companies like Baidu, Alibaba, ByteDance, and DeepSeek, offering unified API, unified authentication, and unified billing.

Operators' platforms internally adapt to multiple large models; users only need to change the model name (Model ID) to switch smoothly.

02 Why Are Operators Selling Tokens?

The explosion of Token service is not accidental.

First, changes in billing models. In the traditional cloud computing era, users were accustomed to paying for "server rental time" or "fixed bandwidth" (i.e., computing power payment at the IaaS layer), buying bandwidth speed and time. However, with the development of large models, the capabilities provided by different models and the costs consumed by different tasks vary greatly. For example, a stronger model costs more per Token; longer contexts consume more Tokens; higher inference complexity leads to higher actual costs. Billing per Token aligns "the level of intelligence consumed by the user" with "the computing power cost paid by the vendor."

Second, lowering technical barriers and "trial-and-error costs." The R&D and deployment of large models often require investments of tens of millions or even billions of dollars. For the vast majority of SMEs and individual developers, building their own models is not realistic. Token service breaks down "Artificial General Intelligence (AGI)" capabilities into pieces, packaging them so developers don't need to worry about how many tens of thousands of GPUs are burning electricity underneath; they only need to call APIs on demand and pay Token fees.

Finally, urgent demand driven by the explosion at the application layer. Entering 2026, application-layer scenarios such as AI Agents, AI-assisted programming, and multimodal content generation have exploded. These applications, in their daily operation, require frequent "throughput" interactions with underlying large models. An automated AI code-writing tool might consume millions of Tokens overnight. This high-frequency, massive-volume interaction forces the market to provide more standardized, stable, and price-competitive Token plan services.

Over the past two decades, operator business models have undergone three core changes in measurement units.

The first stage was the voice era, where operators sold minutes; the second stage was the mobile internet era, selling data GB; and entering the AI era, operators are beginning to experiment with selling Tokens.

Tokens are undergoing a similar evolution to data. Initially, they were just technical metrics; then they became billing units; finally, they evolved into standardized commodities.

The entry of operators marks that Tokens have begun to move beyond the technical realm and enter the consumption system.

In the coming years, the way users purchase AI capabilities may fundamentally change: individual users purchase "AI monthly packages," enterprises procure "Token resource pools," home broadband comes with AI quotas, and government/enterprise dedicated lines integrate Agent services. Tokens will become a basic resource, like electricity, water, and data.

However, this does not mean operators will replace large model vendors.

03 How to Buy Tokens Appropriately?

Should Token service be purchased directly from native large model vendors or from operator platforms? What are the pros and cons of the two current business models?

The first is the native model vendor model, which bills per million Tokens. Vendors like OpenAI, Anthropic, DeepSeek, Qwen, etc., commonly use this system. Users pay separately for input Tokens and output Tokens. Some, like Qwen, might use a pre-purchase at the beginning of the month, settle at the end of the month format.

The second is the operator's monthly subscription Token quota. For example, Shanghai Telecom offers a minimum of 9.9 yuan for 10 million Tokens, with additional purchases for excess, and plans to integrate Token rights into the family's "Beautiful Home" digital space, supporting one-click payment via phone bill.

This "all-in-one price" or "bill integration" model allows Chinese users to purchase large model computing power like they buy data packages.

While overseas markets are dominated by the API tiered pricing of native large model enterprises, the domestic market is pushing Token service into a "packaged" era similar to mobile phone plans.

Currently, both billing models have their advantages, as the user base for Token plans can be divided into three main types.

The first is independent developers and technology enthusiasts (Geeks). They use the API interfaces provided by various vendors to build their own personalized AI applications, such as productivity tools, automatic translation plugins, personal knowledge bases, etc.

The second category is SMEs, startups, and B2B independent software vendors (ISVs). This is the core customer group for Token service. Whether purchasing Tokens for company employees for programming, developing industry-specific AI Agents, or embedding AI assistance into existing enterprise ERP, CRM systems, SMEs need to subscribe to "team edition Token plans" from cloud vendors or operators.

The third category is "AI-heavy dependent" professionals and ordinary households, who need to frequently use AI for copywriting, code writing in home settings, or require AI to tutor their children with homework.

For SMEs and startups, from a techno-economic perspective, the pure Token billing model of native large models is more scientific.

The operator's package model has two advantages. On one hand, independent developers are not tied to one specific large model; they can independently choose from multiple models through the platform provider. On the other hand, Token service may reach mass consumers faster. Most people know what 100GB of data means, but cannot perceive what 10 million Tokens represent.

Operators adopting monthly subscriptions are essentially lowering the cognitive barrier. Users don't need to understand Tokens; they just need to start with a basic package like 9.9 yuan/10 million Tokens to understand their needs.

As operators start selling Token services, "Doubaos" are about to begin competing fiercely at three levels.

From "competing on parameters" to "competing on energy efficiency ratio": For large model companies, they can no longer blindly pursue large parameters and high energy consumption for large models. Instead, they must focus on capabilities like model distillation, quantization, and inference optimization that can output higher quality Tokens with smaller energy consumption.

Price competition will further intensify. After operators aggregate hundreds of models, user switching costs decrease. If model A raises prices, it can be replaced with model B via the platform. When model capability differences are insufficient, price becomes the core competitive factor.

The profit center for large model enterprises will shift. Simply selling APIs offers limited profits. Future profit focus may shift to Agents, industry applications, and enterprise solutions. The model itself gradually becomes infrastructure, while the application layer becomes the value center.

Perhaps, a "two-sided market" is forming: operators control the entry point, model vendors control the capabilities.

Criptomoedas em alta

Perguntas relacionadas

QWhat is the significance of telecom operators launching Token service packages?

AIt signifies that Token is transitioning from a technical metric to a standardized consumer commodity, similar to how mobile data became a utility. It lowers the barrier for AI adoption by offering a familiar, subscription-based model, reduces user lock-in to specific models, and will likely intensify competition among large language model providers.

QHow do the Token purchasing models from native AI firms and telecom operators differ?

ANative AI firms typically charge per million tokens (input/output) with tiered pricing. Telecom operators offer monthly subscription packages with a fixed Token allowance (like 9.9 RMB for 10 million Tokens), bundling it with existing services like phone bills. This model is simpler for general consumers, while the per-token model is more precise for business use.

QWhat are the primary motivations for telecom operators to sell Token services?

AKey motivations include: 1) Adapting the billing model from traditional compute/time to align cost with AI's 'intelligence consumption'. 2) Lowering the technical and trial-and-error barriers for SMEs and developers to access AGI. 3) Meeting the booming demand from AI Agent and other application-layer services that require massive, frequent token interactions. 4) Finding a new core billing unit (like minutes and GB before) for the AI era.

QHow will the emergence of Token platforms impact large language model companies like Doubao?

AIt will weaken user stickiness and force these companies to compete more fiercely ('juan'). Competition will shift to three levels: 1) Efficiency: Improving token quality per unit of energy (via distillation, quantization). 2) Price: Intensifying price competition as users can easily switch models on a platform. 3) Value Shift: Moving their profit center from selling basic API calls to offering higher-value AI Agents, industry applications, and enterprise solutions.

QWhat types of users are the main target audience for Token service packages?

AThree main groups: 1) Independent developers and tech enthusiasts building custom AI tools. 2) SMEs, startups, and B2B software vendors integrating AI into their products or workflows. This is the core audience. 3) Professionals and households that heavily rely on AI for tasks like content creation, coding, or tutoring, who need high-frequency access in daily life.

Leituras Relacionadas

STAR 50 Soars 10.73%, Why Did A-Shares Stage a "V-Shaped Reversal"?

After a prolonged decline, the Chinese A-share market staged a strong rally on July 21. The STAR 50 index surged 10.73%, its largest single-day gain in nearly a year, leading a broad-based "V-shaped" reversal. The Shanghai Composite Index rose 1.79%, the Shenzhen Component Index gained 4.81%, and the ChiNext Index jumped 7.05%. Total market turnover reached 2.97 trillion yuan, an increase of 256.1 billion yuan from the previous session, with over 3,100 stocks advancing. The semiconductor sector spearheaded the rebound, with related ETFs posting significant gains. Analysts attribute the surge to three converging factors. First, coordinated capital inflows from "national team" institutions, insurance funds, listed company buybacks, and fund house self-purchases have bolstered market liquidity and confidence. Second, supportive policy signals, including commitments from regulators to ensure stable market operations, provided a favorable backdrop. Third, a stabilization and recovery in overseas markets, notably South Korea, created a positive external environment. Institutions suggest the most severe panic selling phase for the tech sector has likely passed, following a significant digestion of crowded positions and leveraged funds. While short-term volatility may persist, the medium to long-term outlook remains underpinned by enduring trends like AI computing demand expansion and semiconductor localization. The market's focus now shifts to the sustainability of supportive fund flows, earnings reports, and upcoming catalysts from the global AI industry chain.

marsbitHá 17m

STAR 50 Soars 10.73%, Why Did A-Shares Stage a "V-Shaped Reversal"?

marsbitHá 17m

U.S. Tech Momentum Stocks Post Largest Single-Day Gain Ever, But Is the Plunge Over?

US tech momentum stocks staged a sharp rebound on Tuesday (July 21st). Morgan Stanley's TMT Momentum Factor surged over 12%, marking its largest single-day gain on record, exceeding even peaks from the 2000 dot-com bubble. Key momentum indices from Goldman Sachs also posted their strongest daily performances in years. The rally was led by semiconductors, with the Philadelphia Semiconductor Index jumping 4.6%. This rebound followed three consecutive down days and a cumulative 33% plunge in momentum stocks, one of the steepest drawdowns since the dot-com era. Analysts attribute the surge largely to a short squeeze. Heavy selling had pushed high-beta momentum stocks into deeply oversold territory, forcing many short sellers, particularly in Asia, to cover their positions, creating a self-reinforcing buying spiral. However, the rebound's internals appear weak. Trading volume was notably low, and advancing stocks still lagged decliners on the S&P 500, indicating a narrow, concentrated rally rather than broad market participation. Diverging views emerge on the outlook. BTIG warns the bounce has hit key resistance and recommends selling into strength, citing extreme volatility and historical parallels to past market tops. Conversely, Goldman Sachs and UBS believe the momentum unwind is nearing its end, suggesting it may be time to gradually add exposure, as positioning has been significantly reduced. They caution, however, that high volatility warrants a measured approach, potentially using defined-risk strategies. The upcoming earnings season, particularly reports from major tech firms like Alphabet, is seen as a critical test for the rally's sustainability. Simultaneously, bond markets flashed a warning, with yields rising partly due to spiking oil prices. Analysts note that if long-term Treasury yields break decisively higher, it could pose a significant headwind for equities, especially growth stocks.

marsbitHá 24m

U.S. Tech Momentum Stocks Post Largest Single-Day Gain Ever, But Is the Plunge Over?

marsbitHá 24m

U.S. Tech Momentum Stocks Record Largest Single-Day Gain Ever, but Has the Rout Ended?

U.S. tech momentum stocks staged a dramatic rebound on Tuesday, July 21st. Key momentum indices like the Morgan Stanley TMT Momentum Factor and Goldman Sachs' High Beta Momentum Long Index posted historic or near-historic single-day gains, fueled largely by semiconductor stocks. This sharp rally followed a severe three-day sell-off that saw momentum stocks plunge 33%, marking one of the steepest pullbacks since the dot-com bubble. Analysts attribute the bounce primarily to a short squeeze, as forced covering from over-leveraged traders, particularly in Asia, created a buying spiral. However, the rally's health is questioned due to weak market breadth—overall trading volume was low, and decliners outnumbered advancers in the S&P 500 despite the index's gain—suggesting a narrow, concentrated surge rather than broad recovery. Opinions on the sustainability diverge. BTIG strategists warn the rebound has hit key resistance levels, citing extreme volatility and historic stock dispersion as signs of an ongoing broader correction, and recommend selling into strength. Conversely, Goldman Sachs and UBS view the aggressive momentum unwinding as nearing its end, noting reduced positioning and a lack of new fundamental catalysts. They suggest the sell-off presents a selective opportunity to add exposure, albeit cautiously and gradually using defined-risk strategies. The immediate trajectory hinges on the ongoing earnings season, with market focus on Alphabet's capital expenditure guidance for AI investment clarity. Meanwhile, bond markets present a risk, with rising Treasury yields—potentially heading toward 5.5%—and widening credit spreads for mega-cap tech companies posing a threat to equity valuations. The combination of technical factors, earnings results, and macro conditions leaves the durability of the rebound in doubt.

链捕手Há 27m

U.S. Tech Momentum Stocks Record Largest Single-Day Gain Ever, but Has the Rout Ended?

链捕手Há 27m

Long-Divided Must Unite, Long-United Must Divide: When L1 Becomes Its Own Rollup, What Is Ethereum's Endgame?

"The Inevitable Cycle: When L1 Becomes Its Own Rollup – What is Ethereum's Endgame?" For years, the Ethereum community grappled with concerns that L2s were fragmenting the ecosystem and eroding L1's value. While L2s provided cheaper execution, they also splintered liquidity and the unified user experience of a single chain. This has prompted a fundamental reassessment of the relationship between L1 and L2. Ethereum's roadmap is evolving. The "Scale" initiative merges L1 and L2 expansion into a holistic framework. L1 itself is advancing with higher gas limits, statelessness, and zkEVM verification, no longer content to be just a low-throughput settlement layer. Consequently, the primary value proposition of L2s is shifting from merely providing cheap blockspace to offering L1 cannot easily provide: application-specific optimizations, privacy features, and flexible governance models. L2s are becoming a spectrum of execution environments with varying degrees of security inheritance from Ethereum. A critical challenge in this multi-chain future is interoperability. The vision is to make Ethereum "feel like one chain again." This relies on advancements in native account abstraction (like EIP-7702) and intent-based architectures (Open Intents Framework), where users declare desired outcomes, and solvers handle the complex cross-chain execution. Furthermore, shortening Ethereum's finality time from minutes to seconds is crucial, as it underpins trust between chains for bridges, stablecoins, and cross-chain applications. Perhaps the most provocative idea is that Ethereum L1 itself could become a form of "its own Rollup." As zkEVM and proof systems mature, high-performance nodes could execute transactions and generate validity proofs. Regular validators would then verify these proofs instead of re-executing all transactions. This blurs the traditional L1/L2 hierarchy, making "Rollup" more of a general execution-verification architecture. Native Rollup aims to integrate L2 validation more directly into the Ethereum protocol, allowing L2s to inherit L1's security more fully and move away from reliance on security councils. In the end, L2s are not destined to replace L1 or be made obsolete by it. The likely future is a unified system where diverse execution environments—each optimized for specific use cases like DeFi, gaming, or privacy—coexist. They will share a common foundation of security, liquidity, and verifiable state, seamlessly connected to restore a cohesive user experience. The next phase for Ethereum is not just about scaling through separation, but about intelligently reintegrating what was separated back into a coherent whole.

链捕手Há 43m

Long-Divided Must Unite, Long-United Must Divide: When L1 Becomes Its Own Rollup, What Is Ethereum's Endgame?

链捕手Há 43m

Trading

Spot

Artigos em Destaque

Como comprar ERA

Bem-vindo à HTX.com!Tornámos a compra de Caldera (ERA) simples e conveniente.Segue o nosso guia passo a passo para iniciar a tua jornada no mundo das criptos.Passo 1: cria a tua conta HTXUtiliza o teu e-mail ou número de telefone para te inscreveres numa conta gratuita na HTX.Desfruta de um processo de inscrição sem complicações e desbloqueia todas as funcionalidades.Obter a minha contaPasso 2: vai para Comprar Cripto e escolhe o teu método de pagamentoCartão de crédito/débito: usa o teu visa ou mastercard para comprar Caldera (ERA) instantaneamente.Saldo: usa os fundos da tua conta HTX para transacionar sem problemas.Terceiros: adicionamos métodos de pagamento populares, como Google Pay e Apple Pay, para aumentar a conveniência.P2P: transaciona diretamente com outros utilizadores na HTX.Mercado de balcão (OTC): oferecemos serviços personalizados e taxas de câmbio competitivas para os traders.Passo 3: armazena teu Caldera (ERA)Depois de comprar o teu Caldera (ERA), armazena-o na tua conta HTX.Alternativamente, podes enviá-lo para outro lugar através de transferência blockchain ou usá-lo para transacionar outras criptomoedas.Passo 4: transaciona Caldera (ERA)Transaciona facilmente Caldera (ERA) no mercado à vista da HTX.Acede simplesmente à tua conta, seleciona o teu par de trading, executa as tuas transações e monitoriza em tempo real.Oferecemos uma experiência de fácil utilização tanto para principiantes como para traders experientes.

515 Visualizações TotaisPublicado em {updateTime}Atualizado em 2026.06.02

Como comprar ERA

Discussões

Bem-vindo à Comunidade HTX. Aqui, pode manter-se informado sobre os mais recentes desenvolvimentos da plataforma e obter acesso a análises profissionais de mercado. As opiniões dos utilizadores sobre o preço de ERA (ERA) são apresentadas abaixo.

活动图片