Kimi Forced Two Giants to Change Pricing? Altman Rarely Admits Mistakes, Claude Resets Quotas

marsbitОпубликовано 2026-07-20Обновлено 2026-07-20

Введение

The article discusses the recent competitive dynamics in the AI industry, triggered by the release of Kimi's K3 model and its disruptive pricing. OpenAI's Sam Altman publicly acknowledged past shortcomings while promising an exceptional upcoming year, sparking speculation about GPT-6. Meanwhile, both OpenAI and Anthropic are engaged in a user acquisition battle, aggressively resetting usage limits and offering more credits for their coding agents (Codex/ChatGPT Work and Claude Code) to capture valuable long-task user data. OpenAI's CFO introduced a new efficiency metric: "Useful Intelligence per Dollar," arguing that true cost should be measured by completed work, not just token price. The core competition is shifting from raw model performance to which AI can seamlessly integrate into workflows as a reliable "colleague" that accomplishes tasks, with both companies launching agentic products (ChatGPT Work, Claude Cowork) designed for complex, multi-hour projects.

On the 17th at early morning, the release of Kimi K3 went viral across the internet, marking the second "DeepSeek moment" for open-source AI.

Foreign media Axios reported that K3 is priced far lower than the high-end models it challenges. How long can the high-price strategies of US AI companies last? Coincidentally, around the same time, two giants were engaged in a customer acquisition war. Previously, Altman posted on X, not mentioning a new model, but starting with an admission of fault:

Our performance over the past 12 months has not been ideal, and that's largely my fault.

For someone who has been leading OpenAI in a head-to-head battle with Anthropic, publicly characterizing the past year as "not good enough" and taking the blame is unusual in itself.

What truly shocked the entire internet was the second half of his statement.

Altman then pivoted, stating that OpenAI is about to experience its "best 12 months ever," the team is performing excellently, and they are "seeing results that satisfy them."

What exactly is Altman betting on with "the best 12 months ever"?

The internet's first reaction was almost unanimous: Is GPT-6 coming?

Added 3 Million Active Users

in a Few Days

"Close Twitter and go do your real work!"

Echoing Altman's tweet was OpenAI's Codex lead, Tibo, who dropped a new number of 9 million active users on the 16th.

Looking at the active user numbers stacked together: In February, Codex had less than 1 million; by July 12th, 6 million; July 14th, 8 million; and by the 16th, Codex and ChatGPT Work combined broke 9 million.

Four days, an increase of 3 million people.

Tibo said in his post that he wanted to restore quotas earlier but was held up by the "millions of tasks" the team was busy with to keep the system from crashing and ensure stable operation.

He ended the post reminding everyone that quotas would be restored in a few minutes and to focus on their own work instead of constantly refreshing Twitter.

Simply put, user growth was too rapid, and the engineering team was working around the clock to plug the holes.

This brings to mind Anthropic CEO Dario's "humble brag" a few months ago.

In May, Dario said at a developer conference that the company had originally planned for 10x annual growth, but Q1 revenue and usage, annualized, skyrocketed by 80x: that's where the compute tension came from.

He half-jokingly complained: He really hoped the 80x wouldn't continue, it's too hard to handle, and looked forward to returning to a more normal number, "a mere 10x" would be fine.

Two AI giants, one working non-stop to plug holes, the other wishing its growth would slow down. The fire of agents is burning faster than anyone planned.

In the Same Week

Two Rivals Compete to Give You More Quota

Altman's apology tweet landed right in the middle of a fierce battle of attrition between OpenAI and Anthropic.

Not long after ChatGPT Work's release, OpenAI temporarily removed the 5-hour usage limit for Plus, Pro, and Business plans, and repeatedly reset user quotas: first to about 500,000 users, then rolling out to 7 million users, replenishing everyone once.

Anthropic didn't back down either.

It extended paid access to Claude Fable 5 again and increased the weekly quota for Claude Code by 50%, both valid until July 19th.

One side is removing limits and frantically giving away quota, the other is increasing access quotas and raising limits.

On the surface, both are pampering users; in reality, they are fighting for users.

Looking deeper, this is an arms race in the age of agents: Whoever can retain users during this frenzy will grasp more real long-task data.

The most valuable thing is the real usage data generated when an agent works for you for several hours. Whoever has looser quotas gets more users and usage volume, and thus more data.

And OpenAI's CFO has already applied this calculus to its competitor.

According to OpenAI's statement, on the long-cycle engineering task benchmark DeepSWE v1.1, GPT-5.6 Sol at its highest reasoning tier scored 72.7%, surpassing Claude Fable 5's 69.9%; while estimated API costs are 36.2% lower.

Higher scores, yet cheaper—this is exactly what "how much work per dollar" aims to prove. Simultaneously, output tokens are reduced by 54%, and estimated API costs are lowered by 36.2%.

Interestingly, the named Anthropic did something else almost simultaneously: extended Fable 5 paid access again, increased Claude Code weekly quota by 50%, both until July 19th.

One uses charts to prove it's more cost-effective, the other just opens the floodgates wider for you to try.

This also explains why both companies would rather bear system pressure than push usage volume up. High-intensity real-world usage itself is a moat.

However, some have picked up a different scent from this rapid surge.

Economist Jeremy Nguyen, commenting on "Codex and Claude Code resetting quotas on the same day," said that perhaps many years from now, we'll tell young people about the early "agent token war," how crazy the token subsidies were back then.

He dug up old stories from the dot-com bubble: back then, some startups paid you hundreds of dollars a month just to keep an ad bar on your computer.

His question was pragmatic: If such a window only comes once, how can one make the most of Codex and Claude Code now to get their money's worth?

On this track, Anthropic's Claude Code has been constantly active recently, with enterprise agents for finance and engineering already deployed.

The agent war between the two has just begun.

Hours After the Apology

The CFO Presented a Ledger

Coincidentally, on the same day Altman posted, another article went live on OpenAI's official website.

The author wasn't Altman, but CFO Sarah Friar.

Friar opened by saying that everywhere she goes, CFOs are asking the same question: How to get more value for money from AI. Her answer is a new yardstick—Useful Intelligence per Dollar.

In the past, measuring software success looked at adoption rates: how many seats bought, how many active users, renewal rates. Friar said AI needs a tougher metric: look at how much work is actually accomplished.

She even directly debunked the illusion of "token unit price": Lowest token price ≠ Lowest cost per outcome.

Cheaper model tokens are cheaper, but may require more attempts, more time, and more human review; expensive models get it right the first time. What should really be calculated is the total cost per qualified completed task, the real calculation is the

Full Cost Formula = (Model call cost + Compute resources + Human review time + Retry attempts + Rework cost) ÷ Number of successfully qualified tasks.

What truly determines cost-effectiveness is never the single token price, but the complete cost of one successful task.

The GPT-5.6 family (Sol flagship, Terra balanced, Luna fast and low-cost) was born for this purpose.

A support team's "completion" is a customer issue resolved; an engineering team's "completion" is a code change passing tests; a legal team's "completion" is a contract reviewed without errors.

Reading this and looking back at the earlier quota war, the flavor changes.

OpenAI removing the 5-hour limit and repeatedly resetting quotas, Anthropic increasing Claude Code weekly quota by 50%—this isn't just subsidizing users to grab data. When the unit of measurement shifts from "seats" and "tokens" to "completed work," giving away quota isn't a concession, it's changing the metric.

Whoever gets everyone used to calculating based on "how much work is done" first redefines how this market competes.

Capability wins the first use, reliability makes AI become the work process itself.

As usage grows, is every dollar of AI creating more value for you?

The answer lies in the compute infrastructure flywheel: Better infrastructure → Stronger models → Better products → Higher adoption → More revenue → Sustained investment in next-gen research and compute.

When, over time within the same workflow, the number of successful tasks grows faster than the total cost, and quality improves rather than declines, "useful intelligence per dollar" achieves positive compounding.

What the Duopoly is Fighting Over

is Who Becomes Your AI Colleague

Taking a broader perspective makes things clearer.

Anthropic struck first earlier this year. Claude Code became legendary among developers, and the subsequent Cowork pushed agents from programmers to general knowledge workers, taking the lead in "AI office work."

OpenAI's ChatGPT Work, released on July 9th, almost directly targeted Cowork's selling points, built-in Codex, and even launched the latest model GPT-5.6 Sol released the same day.

A command like "Help me create a project tracking table" results in a Gantt chart with 18 projects and 29 tasks: This isn't chatting, it's delivering work.

What changed this time isn't parameters, it's form.

In the past, using ChatGPT meant you asked, it answered; now you just give one goal, and it handles the rest: breaking down tasks, calling tools, searching your files and apps, working for hours straight until the job is actually done.

The official statement clarifies the positioning: ChatGPT is no longer just a machine that answers questions, but a partner for tackling complex work.

OpenAI also experimented on its own company: Today, nearly 100% of internal teams, from finance to sales, use ChatGPT Work and Codex to get work done.

Furthermore, compared to Anthropic, OpenAI holds a card that's hard to beat short-term: distribution.

Cowork requires users to actively download a desktop client; but integrating agents into ChatGPT means over 900 million weekly active users are already using agents the moment they open that familiar app.

Altman himself stated that the usage of agentic products increased 2.5x week-over-week. This curve is part of the foundation for his "strongest next year" confidence.

Returning to Altman's prediction.

He's likely betting not on a new model called GPT-6, but on a more fundamental form shift: transitioning AI from "answering questions" to officially "doing work for you." And Friar's yardstick is precisely prepared for this shift—when AI starts doing work, what measures it is no longer parameters, but how much work is done and how much it costs.

One makes statements on the front stage, the other changes the accounting ledger backstage.

The final outcome of this competition may not depend on whose model is stronger, but on who truly moves into everyone's workflow first, becoming your AI "colleague."

References:

https://x.com/sama/status/2077817060068057493?s=20

https://x.com/JeremyNguyenPhD/status/2077719990116258162

This article is from the WeChat public account "New Zhiyuan", author: ASI Apocalypse, editor: Yuanyu Aeneas

Связанные с этим вопросы

QWhat is the main focus of OpenAI's new metric, 'Useful Intelligence per Dollar', and why is it significant?

AThe main focus of OpenAI's new metric, 'Useful Intelligence per Dollar', is to measure the cost efficiency of AI by calculating the total cost required to successfully complete a qualified task. It is significant because it shifts the evaluation paradigm from token price and adoption rates to the actual value and productivity delivered by AI. It factors in model calls, compute resources, human review time, retries, and rework costs. This redefines market competition, emphasizing real-world work completion over simple usage or cheap tokens.

QWhat triggered the recent 'token war' or 'allocation battle' between OpenAI and Anthropic, and what are both companies trying to achieve with it?

AThe battle was triggered by the rapid growth and competitive pressure in the 'agent' AI market, exemplified by launches like Anthropic's Claude Code/Cowork and OpenAI's ChatGPT Work with Codex. Both companies are trying to capture and retain users by offering more generous usage limits and resetting allowances. Their deeper goal is to acquire high-quality, real-world usage data from long-duration tasks performed by their AI agents. The company that secures more of this valuable data can improve its models faster and potentially dominate the emerging workflow integration market.

QAccording to the article, what fundamental shift in AI product philosophy do the latest releases from OpenAI and Anthropic represent?

AThe latest releases represent a fundamental shift from AI as a tool for 'answering questions' to AI as a 'partner that does work.' Products like ChatGPT Work and Claude Cowork are designed to autonomously handle complex, multi-step tasks over hours. Users provide a goal, and the AI agent breaks it down, uses tools, accesses files, and executes until the work is genuinely completed. This changes the AI's role from an interactive assistant to an integrated, proactive colleague within a user's workflow.

QHow does the article characterize Sam Altman's rare public admission of fault, and what does it connect this to?

AThe article characterizes Sam Altman's public admission that 'the last 12 months have not been great for OpenAI and that's mostly my fault' as highly unusual and strategic. It connects this statement to his immediate follow-up prediction of 'the best 12 months in OpenAI's history.' The article suggests this 'apology-preview' combo is not about a single model like GPT-6, but about betting the company's future on successfully transitioning users to the new 'agentic' AI paradigm where AI becomes a core work partner, a shift he believes will define OpenAI's coming year.

QWhat competitive advantage does OpenAI's ChatGPT Work have over Anthropic's Claude Cowork regarding user distribution, according to the analysis?

AAccording to the analysis, OpenAI's key distribution advantage is its massive existing user base. Integrating the smart agent functionality directly into the ChatGPT app means potentially over 900 million weekly active users are instantly exposed to the new 'Work' features. In contrast, Anthropic's Claude Cowork requires users to actively download a separate desktop application. This gives OpenAI a significant head start in user acquisition and habituation, lowering the barrier for users to try and adopt the agentic workflow.

Похожее

Патрик Витт откладывает военную подготовку, чтобы возглавить усилия Белого дома по принятию закона CLARITY Act

Белый дом усилил работу над Законом о CLARITY, и его главный советник по цифровым активам, Патрик Уитт, отложил свою военную подготовку в Национальной гвардии армии Джорджии, чтобы продолжить переговоры по этому важному законопроекту о криптовалютах. Уитт должен был начать обучение 27 июля, но из-за предстоящего перерыва в работе Сената с 8 августа открылось небольшое окно для переговоров. Администрация считает этот период критически важным для продвижения Закона о CLARITY, который призван установить четкие правила регулирования рынка цифровых активов в США. Это уже не первый раз, когда Уитт откладывает военную службу ради этой работы. Актуальность роли Уитта возросла после того, как заместитель директора Гарри Джунг объявил о своем уходе из правительства. Первоначально Джунг должен был взять на себя обязанности во время отсутствия Уитта, но теперь присутствие Уитта в Вашингтоне необходимо для обеспечения преемственности руководства в ключевой законодательный период. Администрация продолжает работу над законопроектом до перерыва в Сенате 8 августа.

TheNewsCrypto3 мин. назад

Патрик Витт откладывает военную подготовку, чтобы возглавить усилия Белого дома по принятию закона CLARITY Act

TheNewsCrypto3 мин. назад

Акции Circle обвалились на 76%, гонконгская стейблкоин появится в течение двух недель

Акции Circle упали на 76% с июня 2023 года, с $260 до $62, что свидетельствует о переоценке рынком ее стоимости. Президент компании Хит Тарберт заявляет о долгосрочной стратегии, но аналитики Mizuho понизили рейтинг акций, ссылаясь на растущую конкуренцию и давление на прибыль. Конкуренция на рынке стейблкоинов обостряется. USDC сохраняет лидерство с объемом в $730 млрд, но сталкивается с вызовом от нового стейблкоина Open USD и платформы Visa Stablecoin Platform, которые используют модель распределения доходов для привлечения партнеров. Circle в ответ развивает партнерство с японской JCB для интеграции USDC в офлайн-платежи. Tether (USDT) сталкивается с проблемами соблюдения требований нового закона США (GENIUS), дающего два года на приведение резервов (сейчас включающих биткойны и драгметаллы) в соответствие с нормами. Несоблюдение может ограничить его использование на рынке США. В Гонконге ожидается запуск стейблкоина HKDAP, привязанного к гонконгскому доллару, после получения одной из первых лицензий регулятором консорциумом во главе со Standard Chartered. Ключевым вопросом станет интеграция в реальные финансовые потоки. Падение акций Circle отражает не только усиление конкуренции в отрасли, но и изменение рыночных ожиданий относительно процентных ставок. Эпоха безраздельного доминирования на рынке стейблкоинов подходит к концу, и ценность масштаба пересматривается. Будущее определит, кто сможет реализовать заявленные стратегии.

marsbit7 мин. назад

Акции Circle обвалились на 76%, гонконгская стейблкоин появится в течение двух недель

marsbit7 мин. назад

В условиях капиталистической осады децентрализация — единственная линия обороны публичных блокчейнов

В статье профессора Колумбийской бизнес-школы Омида Малекана утверждается, что в условиях давления со стороны крупного капитала и корпораций единственной надежной защитой публичных блокчейнов является максимальная децентрализация. Автор, называющий себя прагматиком и реалистом, анализирует ситуацию с позиций макиавеллизма и истории, доказывая, что любые компромиссные и централизованные решения (разрешенные блокчейны, консорциумы) неизбежно будут поглощены или коррумпированы традиционными институтами для сохранения их власти и прибыли. Главная опасность для протоколов заключается не во внешних атаках, а в захвате контроля изнутри, о чем свидетельствует эволюция таких платформ, как Visa, Mastercard или Google. Поскольку публичный базовый блокчейн потенциально охватывает огромный рынок расчетов для активов, платежей и приложений, мотивация для его захвата чрезвычайно высока. Корпорации будут либо пытаться контролировать новые технологии, либо активно их дискредитировать. Псевдодецентрализованные корпоративные решения, по мнению автора, неэффективны и служат лишь тактикой задержки, позволяя традиционным игрокам замедлить внедрение truly децентрализованных систем. Однако в долгосрочной перспективе рынок, подобно воде, стекающей вниз, неизбежно выберет наиболее безопасную и открытую инфраструктуру, несмотря на ее текущие недостатки и высокую стоимость работы. Ethereum, при всех своих несовершенствах, сегодня представляет собой оптимальный баланс и наилучшую защиту от захвата.

Foresight News17 мин. назад

В условиях капиталистической осады децентрализация — единственная линия обороны публичных блокчейнов

Foresight News17 мин. назад

Кто решает правила Биткоина? BIP-110 вызывает раскол в управлении

**Резюме: BIP-110 и спор о том, кто определяет правила Биткоина** Предложение BIP-110 (Reduced Data Temporary Softfork) вызвало глубокие разногласия в сообществе Биткоина, выйдя за рамки технических дебатов и подняв фундаментальный вопрос управления: кто имеет право решать, что такое Биткоин и каковы его правила? **Суть BIP-110:** Предложение направлено на временное (на ~1 год) ограничение некриптовалютных данных (например, из Ordinals, Runes) непосредственно на уровне консенсусных правил, делая некоторые в настоящее время допустимые транзакции недействительными. Это повысило бы стоимость записи произвольных данных в блокчейн. Критики, включая Майкла Сайлора и Адама Бэка, утверждают, что это опасный прецедент, подрывающий нейтральность и устойчивость протокола. **Ключевые линии конфликта:** 1. **Уровень изменений:** BIP-110 пытается перенести борьбу с "спамом" с уровня политики ретрансляции/майнинга (который, по мнению сторонников, уже неэффективен из-за обходных путей) на уровень базовых консенсусных правил. 2. **Легитимность изменений:** Стороны спорят о легитимном механизме принятия решений: * **Майнеры:** Некоторые (например, F2Pool) считают, что PoW — это "конституция", и решения требуют поддержки майнеров. * **Ноды:** Сторонники Bitcoin Knots настаивают на суверенитете операторов нод, которые бесплатно обеспечивают безопасность. * **Разработчики:** Команда Bitcoin Core, изменившая политику по умолчанию в v30, обладает влиянием через код, но не имеет формального мандата. * **Крупные холдеры:** Майкл Сайлор, представляющий интересы корпоративных казначейств, вносит в спор вес нарратива и капитала. * **Технический консенсус:** Адам Бэк отстаивает модель, при которой любое изменение требует тщательного, медленного обсуждения среди разработчиков, что служит "иммунной системой" протокола. 3. **Практические проблемы:** Даже если BIP-110 активируется, он, вероятно, не сможет полностью остановить запись данных, так как всегда найдутся обходные методы. Кроме того, в реализации была обнаружена потенциальная уязвимость (BlockSlop), которая могла бы привести к расхождениям между нодами. **Итог:** BIP-110 стал стресс-тестом для системы управления Биткоином, где нет единого центра власти. Конфликт выявил противоречия между различными группами (майнеры, ноды, разработчики, крупные инвесторы), каждая из которых апеллирует к своему источнику легитимности. Вопрос о том, какая из этих сил в конечном счете определяет эволюцию протокола, остаётся открытым.

marsbit32 мин. назад

Кто решает правила Биткоина? BIP-110 вызывает раскол в управлении

marsbit32 мин. назад

Кто решает правила биткойна? BIP-110 обостряет разногласия по управлению

BIP-110 «Снижение данных временным софтфорком» вновь расколол сообщество Bitcoin по вопросу о том, кто определяет правила сети. Предложение, направленное на ограничение нефинансовых данных (например, из Ordinals и Runes) через консенсусные правила, а не политику ретрансляции, вызвало широкие дебаты. Ключевые стороны: основатель MicroStrategy Майкл Сэйлор выступил против, приведя 110 доводов, включая низкий 55% порог активации и риски раскола цепи. Сооснователь Blockstream Адам Бэк подчеркнул, что технический консенсус и «разрешительный» характер Bitcoin — его иммунная система. Разработчики Bitcoin Core v30, смягчившие политики OP_RETURN, считают, что ноды утратили эффективный контроль над данными. Майнеры разделились: Ocean поддержал BIP-110, Foundry проводит голосование клиентов, а F2Pool высказался против. Обнаруженная уязвимость BlockSlop в клиенте BIP-110 высветила технические риски. Критики указывают, что предложение может не остановить запись данных, а лишь поднимет её стоимость, и создаст опасный прецедент в управлении. Спор выявил глубокий раскол между майнерами, разработчиками, операторами нод и крупными держателями BTC, такими как MicroStrategy, о легитимности и будущем Bitcoin. В итоге, BIP-110 стал стресс-тестом управления, поставив главный вопрос: кто решает, чем должен быть Bitcoin.

链捕手44 мин. назад

Кто решает правила биткойна? BIP-110 обостряет разногласия по управлению

链捕手44 мин. назад

Торговля

Спот
活动图片