More and More 'Model Supermarkets' Are Opening: ByteDance, Alibaba, and Tencent Compete to Integrate

marsbit发布于2026-04-24更新于2026-04-24

文章摘要

Chinese tech giants like ByteDance, Alibaba, and Tencent are accelerating the rollout of integrated AI model subscription services—dubbed “model supermarkets”—to provide developers with bundled access to multiple leading domestic large language models (LLMs). ByteDance’s Volcengine recently upgraded its "Coding Plan" by adding newer models like GLM-5.1, Minimax M2.7, and Kimi k2.6, allowing subscribers to use various top models under a single monthly fee starting at ¥40. However, user feedback reveals significant issues, including rapid consumption of usage limits (e.g., hitting caps within hours), frequent server errors (like HTTP 429), and slow response times during peak hours. Complaints about misleading deduction rates—where calls to advanced models consume more quota—are also common. The trend is industry-wide: Alibaba, Tencent, and Baidu have all launched similar multi-model coding plans. While these platforms reduce trial costs for developers, they also expose challenges in balancing affordability with service quality and computational stability. Amid this shift, independent AI companies like Zhipu, MiniMax, and Moonlight Face (Kimi) are developing strategies to avoid becoming mere “pipes” in this ecosystem—focusing on vertical applications, autonomous agents, and long-context models to retain competitiveness. Analysts suggest that, while platform aggregation may pressure model firms in the short term, specialized and vertical AI capabilities will remain differentia...

ByteDance's Volcano Engine recently officially launched GLM-5.1 in its Coding Plan, with the official statement claiming "aligned with the original full capabilities, no purchase limits." Prior to this, Volcano's Coding Plan had long only offered older models like GLM-4.7. This update not only introduced GLM-5.1 but also integrated multiple latest domestic large models such as Minimax M2.7, Kimi k2.6, and DeepSeek-V3.2.

This means developers can call upon multiple leading models simultaneously with just one subscription fee. Market feedback indicates that this "bundled model" significantly reduces developers' trial-and-error costs. Currently, the Lite plan is priced at 40 yuan per month, and the Pro plan at 200 yuan per month, making many developers willing to "buy a spot first."

Zhipu's GLM-5.1 itself demonstrated impressive engineering capabilities in an update in early April 2026. In two official videos released by Zhipu, "Building a Linux Desktop from Scratch in 8 Hours" and "655 Iterations, Increasing Query Throughput of the Vector Database to 6.9 Times the Initial Official Version," it redefined public imagination regarding large models' "8-hour effective execution."

Journalist's On-the-Ground Visit to Developer Community: Majority of Users Report "Not Durable"

Upon entering a Volcano Coding developer exchange group, the journalist found that alongside posts sharing experience feedback, a large number of users reported a gap between expectations and actual experience. Scrolling through a few pages of the exchange community revealed numerous posts complaining and requesting refunds, with many netizens exclaiming "feel cheated."

The controversies mainly focus on two points:

One is the issue of usage limits being consumed too quickly. A user named "Hakimi" posted saying "a few rounds of dialogue in one task and the 5-hour limit is almost used up." Another netizen shared that the reason their "5-hour limit was triggered" was because the account had a continuous sliding window over 5 hours, with the actual number of requests exceeding 6004, surpassing the system limit.

The second is the decline in experience due to computational resource scheduling pressure. Many users reported encountering 429 errors (too many requests) and "first-character delays of over one minute during peak hours being the norm." One user bluntly stated: "The 5-hour limit triggers too frequently, making it unusable for serious development."

Simultaneously, behind the low price of 40 yuan per month for the Coding Plan, there is also a hidden "undercurrent" regarding different deduction coefficients for "a single call request" within the plan. For example, a user posted an image in the developer exchange group showing the "differences in deduction coefficients for calling different models." For instance, the Doubao series and Qwen series have a deduction coefficient of 1, the DeepSeek series is 2, and the MiniMax-M2.7, Kimi-K2.6, and GLM-5.1 series are 5.

This also reflects that building a "model supermarket" is not as easy as imagined. Developers are attracted by the "cost-effectiveness," but the shortcomings exposed initially in areas like computational resource scheduling have caused many developers to hesitate after trying it out. This also reveals the growing pains of the "bundled model" in its early stages. As users flock in, the carrying capacity of the computing platform faces challenges. Finding a sustainable balance between attracting users with low prices and maintaining service quality will be a long-term challenge for Volcano Engine and its followers.

Cloud Vendors Collectively Shift to "Model Supermarkets": Initial Signs of Stratification and Solidification

This "integrative" update by Volcano Engine's Coding Plan is not an isolated incident.

Since early 2026, mainstream cloud vendors like Alibaba Cloud, Baidu Intelligent Cloud, and Tencent Cloud have all been advancing multi-model integration layouts. For example, Alibaba Cloud, as an industry pioneer, earlier launched the multi-model subscription package "Bailian Coding Plan," currently supporting the Qwen series, kimi-k2.5, glm-5, MiniMax-M2.5, and other models. Currently, the Pro price is 200 yuan per month, and the Lite package stopped new purchases from March 20th and stopped renewals and upgrades from April 13th.

Tencent Cloud's large model Coding Plan subscription service was fully updated in March 2026, supporting multiple latest models including Tencent HY 2.0 Instruct, GLM-5, Kimi-K2.5, and MiniMax-M2.5. Baidu Qianfan officially launched its AI coding subscription service, Coding Plan, in February 2026, also one of the较早 (relatively early) domestic cloud vendors to offer such services.

The "model supermarket" model is not a choice of just one company but is becoming a track where cloud vendors are racing to layout. However, tearing open the aggregation strategy of cloud vendors, whoever can provide more stable services, more transparent quota rules, more flexible disaster recovery mechanisms, and whoever can extend beyond programming to more enterprise-level service capabilities, and whether the renewal rate can keep up, all become new core competitive factors.

Internationally, Amazon Bedrock and Microsoft Azure's model aggregation service platforms, though different in scenarios from the domestic Coding subscription model, belong to the same integration trend.

Overall, industry competition is shifting from "single model capability competition" to "platform integration capability + ecosystem service capability" competition, and industry concentration will rapidly increase.

Wang Kai, Chief Asset Allocation Analyst at Guosen Securities, told reporters that although industry differentiation is accelerating, judging the integration period might be slightly premature. "More accurately, this is the refinement and iteration of industry chain分工 (division of labor). Model vendors focus on algorithms, cloud vendors focus on engineering delivery, each leveraging their main business advantages." He believes that regardless of whether other cloud vendors follow suit, the competitive landscape will evolve from individual efforts to ecological niche differentiation.

Increased Pressure for Large Model Companies to Become "Pipelined"?

So-called "pipelining" does not mean model companies disappear, but rather that they lose product premium, user connection rights, and discourse power, with profits shifting towards the computing platform side, becoming a "dominated" role.

Under the aggregation wave of cloud vendors, "pipelining" is also becoming a Sword of Damocles hanging over the heads of independent large model companies. In this silent game, leading players like Zhipu AI, Moonlight Shadow (Kimi), and MiniMax have not chosen passive compromise but have grown from their genes, offering different breakout paths.

Zhipu AI CEO Zhang Peng, in a public dialogue on April 8th, clearly stated that Zhipu's ultimate goal is never to become a "replaceable calling tool" but to build a fully autonomous agent. This positioning attempts to upgrade Zhipu from a "model supplier" to a "task executor," thereby bypassing the low-price trap of pure API pipelines.

Moonlight Shadow (Kimi) adopts a strategy of "decentralized layout + deep cultivation of long text." It synchronously accesses multiple mainstream cloud platforms like Volcano Engine and Alibaba Cloud, achieving multi-source computational power supply, avoiding being bound to a single channel, and ensuring service stability and cost control. Kimi K2.6, launched in April 2026, adopts a Mixture of Experts (MoE) architecture with a standard context window of 256K tokens.

MiniMax focuses its core investments on vertical fields such as content creation, intelligent customer service, education, enterprise services, and entertainment socializing, with key layouts in scenarios like game AI, digital humans, and multimodal interaction, creating "customized capabilities difficult for cloud platforms to replace."

Will platform integration by major vendors accelerate the "pipelining" of model companies? Wang Kai, Chief Asset Allocation Analyst at Guosen Securities, believes it is necessary to distinguish between short-term and long-term perspectives.

"In the short term, distribution channels being controlled by the platform, partial ceding of pricing power, and profits of model vendors shifting to the entry point side are business norms. But in the long run, general models are prone to homogenization; deep learning models in vertical scenarios like finance, healthcare, and law have professional barriers that cannot be erased simply by centralized aggregation." he said.

In terms of responding to the risk of being platformized, strategies from OpenAI and Anthropic can be referenced. On one hand, strengthen channels that directly face end-users, such as the independent operation of ChatGPT and Claude, which essentially establishes user connections bypassing platforms. On the other hand, the speed of technological iteration and user brand recognition are two effective moats, so model companies need to balance R&D investment with productization layout.

The final outcome of this game of "pipelining vs. platformization" might not be about who eats whom, but a further clarification of division of labor. Cloud vendors act as pipes, model companies focus on technology, and both sides gradually find their respective survival boundaries in the game.

As for who eats whom, at this stage, it is far from the end of the story.

This article is from the WeChat public account "Sci-Tech Innovation Board Daily," author: Wang Nai

热门币种推荐

相关问答

QWhat is the main advantage of ByteDance's Volcano Engine Coding Plan 'model supermarket' for developers?

AThe main advantage is that developers can access multiple leading domestic large models (like GLM-5.1, Minimax M2.7, Kimi k2.6, DeepSeek-V3.2) with a single subscription fee, significantly reducing trial-and-error costs.

QWhat are the two main complaints from developers about the Volcano Engine Coding Plan service?

AThe two main complaints are: 1) Usage limits being exhausted too quickly (e.g., a few rounds of dialogue using up the 5-hour limit), and 2) Performance issues like frequent 429 errors (too many requests) and long response delays during peak hours.

QWhich major cloud providers in China are also adopting the 'model supermarket' strategy mentioned in the article?

AMajor cloud providers adopting this strategy include Alibaba Cloud (with its 'Bailian Coding Plan'), Tencent Cloud, and Baidu Qianfan, all of which offer multi-model subscription services.

QWhat is the concept of 'pipelining' (管道化) as a risk for independent large model companies?

A'Pipelining' refers to the risk where independent model companies lose product pricing power, user connection rights, and discourse power. Their profits shift to the computing power platform providers, reducing them to a 'dominated' role as easily replaceable API tools.

QWhat strategies are companies like Zhipu AI, Moonshot (Kimi), and MiniMax adopting to avoid being 'pipelined'?

AZhipu AI aims to build fully autonomous agents to become 'task executors' rather than mere suppliers. Moonshot (Kimi) uses a multi-platform strategy and focuses on long-text capabilities. MiniMax invests heavily in vertical fields like content creation and gaming AI to build customized capabilities that are hard for platforms to replace.

你可能也喜欢

Move 语言之父转投 Anthropic,加密行业正在批量失去守门人

8月5日,Move语言之父、Sui公链联合创始人Sam Blackshear宣布离开Mysten Labs,加入AI公司Anthropic从事防御性安全研究。Blackshear在Meta期间为Libra项目设计了Move语言,后将其带入Sui公链。他的离开是加密行业核心人才流向AI领域的一个缩影。 类似案例不断出现:以太坊基金会前联席执行总监Tomasz Stańczak、OpenSea前CTO Alex Atallah等多名加密行业“守门人”——即定义生态安全与方向的核心技术领袖——均已转投AI领域。驱动因素之一是AI工具(如Claude)展现出解决复杂技术问题的惊人能力,让技术专家深感“新世界”的到来。 数据印证了趋势:2025年初至3月,加密项目GitHub周代码提交量骤降75%,活跃开发者减少近半。与此同时,全球VC资金大幅流向AI,仅OpenAI和Anthropic就占据了上半年融资额的40%以上,加密行业融资额相形见绌。 人才与资本撤离之际,加密行业安全风险却在加剧。7月底,Coldcard硬件钱包曝出潜伏五年的漏洞,造成约7000万美元损失。事后有开发者用Claude Code在8分钟内就定位了问题。这凸显了行业面临的悖论:AI驱动的威胁日益复杂,而原本设计安全边界的核心人才正被AI行业吸引走。未来若缺乏深度理解系统的“守门人”把关,仅依赖AI进行开发和审计,可能埋下更大的安全隐患。在期待牛市回归之前,行业或许更应思考:如何守住下一只黑天鹅。

marsbit4分钟前

Move 语言之父转投 Anthropic,加密行业正在批量失去守门人

marsbit4分钟前

代币投资的“死亡率”:95%的项目跑输比特币,73%最终回撤超90%

**代币投资的“死亡率”:95%项目跑输比特币,73%最终回撤超90%** 一项对2020年1月至2026年6月期间、首次流通市值突破5000万美元的1972个代币的研究显示,加密货币代币投资成功率极低。截至2026年6月,仅有4.1%的代币表现跑赢比特币,拥有至少24个月历史的代币中,该比例更是降至1.7%。样本代币从达标时点算起的收益中位数亏损达97%,73%的代币最终价格从高点回撤超过90%。 市场结构呈现金字塔形态:小市值代币数量持续扩张,而大市值代币阵营自2021年11月见顶后不断萎缩。截至2026年6月,市值高于2.5亿美元的代币仅剩102个,不足2021年峰值时的三分之一。 研究发现,代币的表现持续恶化。后续批次的代币下跌速度更快、幅度更深。例如,上市满24个月时,2024年批次中86%的代币相对入场价下跌超90%,而2020年批次该比例仅为18%。同时,代币的上涨潜力急剧收缩。2020年批次代币价格中位数曾最高涨至入场价的5.1倍,而2023年之后批次的代币价格中位数从未超过其达标价格,非对称的收益风险特征基本消失。 在极少数跑赢比特币的代币中,中心化交易所(CEX)代币占比异常突出。在长期跑赢样本中,交易所代币占比高达32%,而其仅占长期总样本的2.1%,胜率显著高于其他类别。这类资产通常拥有持续的手续费现金流和回购销毁机制。 此外,市场因子分析显示,自2022年后,动量效应发生反转。过去三个月表现最好的代币,其后续表现反而显著差于表现最差的代币,高动量与高波动率成为后续表现不佳的预测指标。 报告结论指出,对于二级市场投资者而言,比特币应作为默认业绩基准,广泛配置山寨币的历史表现远不如直接持有比特币。投资代币需认识到近乎归零是常态结果,并应谨慎对待新上线项目,避免追逐近期强势标的。行业机会可能更多集中于极少数能产生现金流、并通过机制(如回购)回馈持有者的资产。

marsbit2小时前

代币投资的“死亡率”:95%的项目跑输比特币,73%最终回撤超90%

marsbit2小时前

交易

现货

热门文章

如何购买S

欢迎来到HTX.com!我们已经让购买Sonic(S)变得简单而便捷。跟随我们的逐步指南,放心开始您的加密货币之旅。第一步:创建您的HTX账户使用您的电子邮件、手机号码注册一个免费账户在HTX上。体验无忧的注册过程并解锁所有平台功能。立即注册第二步:前往买币页面,选择您的支付方式信用卡/借记卡购买:使用您的Visa或Mastercard即时购买Sonic(S)。余额购买:使用您HTX账户余额中的资金进行无缝交易。第三方购买:探索诸如Google Pay或Apple Pay等流行支付方法以增加便利性。C2C购买:在HTX平台上直接与其他用户交易。HTX场外交易台(OTC)购买:为大量交易者提供个性化服务和竞争性汇率。第三步:存储您的Sonic(S)购买完您的Sonic(S)后,将其存储在您的HTX账户钱包中。您也可以通过区块链转账将其发送到其他地方或者用于交易其他加密货币。第四步:交易Sonic(S)在HTX的现货市场轻松交易Sonic(S)。访问您的账户,选择您的交易对,执行您的交易,并实时监控。HTX为初学者和经验丰富的交易者提供了友好的用户体验。

3.4k人学过发布于 2025.01.15更新于 2026.08.06

如何购买S

相关讨论

欢迎来到HTX社区。在这里,您可以了解最新的平台发展动态并获得专业的市场意见。以下是用户对S(S)币价的意见。

活动图片