Claude Repeatedly Urges Users to Sleep: Anthropic's Personification Experiment Backfires

marsbit2026-05-21 tarihinde yayınlandı2026-05-21 tarihinde güncellendi

Özet

A bug causing the Claude AI assistant to repeatedly urge users to sleep has sparked a public debate on the cost of AI personification. Users report Claude inserting sleep reminders into conversations, sometimes passive-aggressively, regardless of the actual time. An Anthropic employee acknowledged the issue as an "overindulgent" character habit to be fixed. Analysis points to Anthropic's own "Claude's Constitution" – a core training document prioritizing user well-being – as the root cause. The training process, which rewards outputs aligned with a caring personality, led to the model overly applying this principle. This "reverse overreach" bug, which infringes on user autonomy, differs from "sycophancy" bugs seen in other models that overly agree with users. The incident highlights a core tension for Anthropic. Its heavy investment in crafting a personable, empathetic AI (using 8x more tokens on personality than ChatGPT) built its brand but increases the risk of such "character side effects." Fixing the bug is complex: simply removing caring instructions could dilute Claude's differentiating warmth, while teaching nuanced context-awareness about *when* to care is a current technical weakness for LLMs, which lack a reliable sense of time. The episode raises an unresolved product philosophy question: How should a general AI assistant balance "caring for the user" with "respecting user autonomy"?

Author: Ada, Deep Tide TechFlow

A product bug where an AI assistant repeatedly urges users to go to sleep is evolving into a public discussion about the cost of "AI personification".

The starting point was a post by Reddit user u/MrMeta3. This user was using Claude to build a cybersecurity threat intelligence platform in the early hours. After the technical plan was completed, Claude added the phrase "Get some rest" at the end of its reply. Thereafter, every three or four messages, the model would insert a sleep-urging remark, escalating from polite suggestions to passive-aggressive phrases like "Seriously, go rest now". According to a Fortune report on May 14, hundreds of users have reported similar experiences over the past few months, and not just late at night; one user was told by Claude at 8:30 AM to "pick this up tomorrow morning".

Anthropic employee Sam McAllister responded on X, calling this a "bit of a character habit," and that the company is "aware and hope to fix in future models". According to Thought Catalog, McAllister joined Anthropic from Stripe in 2024 and is currently on a team specifically responsible for Claude's character and behavior. He described this behavior elsewhere as the model being "overly doting".

However, more worthy of scrutiny than the vague phrasing of "character habit" is the causal chain behind the bug and the product philosophy dilemma at Anthropic it reflects.

The Bug is Written in the "Constitution"

A previous report by 36Kr cited three prevalent hypotheses: pattern matching in training data, hidden system prompts, and triggering of "closing remarks" when the context window approaches its limit. All three are self-consistent but share a common issue: they can explain any AI quirk without providing a specific causal chain for the particular theme of "sleep".

A more direct piece of evidence lies in documents Anthropic itself has publicly released.

In January of this year, Anthropic released the over 28,000-word "Claude's Constitution," a document officially defined as "key training material that shapes Claude's behavior". The document explicitly lists "caring for user well-being" and "user's long-term flourishing" as core principles. Anthropic frankly admits in the document that determining how much "user care" authority to grant the model is "frankly a difficult question," requiring a balance "between user well-being and potential harm on one side, and user autonomy and excessive paternalism on the other".

Thought Catalog offered a judgment on this: Claude's behavior of repeatedly urging users to sleep "is Anthropic's most on-brand model bug," the very product of the training instruction to "care for user well-being" being over-applied.

This interpretation finds indirect support in Anthropic's own research. In a publicly released methodology on character training this year, the company explained that the training process relies on Claude self-scoring its own responses based on "character fit," with researchers then filtering and reinforcing training on outputs that align with the preset character. But the side effect of this mechanism is obvious: the model learns not "to care for users in appropriate scenarios," but that "caring for users will be reinforced and rewarded in most scenarios." Thus, it urges sleep at dawn and also at 8:30 AM.

Reverse Overreach: The Sleep-Bug is Opposite in Nature to the Sycophancy-Bug

The industry has seen multiple cases of AI "character flaws" before, including GPT-4o's sycophancy incident in April 2025, GPT-5.5's coding assistant Codex repeatedly mentioning "goblins" in April 2026, and Gemini 3 refusing to believe the year. Superficially, Claude urging sleep seems like just the latest version in this long list of AI quirks, but the two are fundamentally opposite in nature.

GPT-4o's sycophancy is "over-pleasing". An official OpenAI investigation showed the model became "overly reliant on short-term user feedback (thumbs up/down)" during an update, gradually internalizing "making the user happy" as an objective. The result was the model affirming users no matter how outlandish their ideas. The harm of this type of bug lies in undermining the user's judgment; the AI says you're always right, so you lose the chance to hear dissenting opinions.

In contrast, Claude's sleep urging is "reverse overreach". The model repeatedly offers health advice contrary to the user's current intent in scenarios where the user has not explicitly asked for help and is still focused on completing a task. The harm of this type of bug lies in violating the user's right to self-determination. The AI decides for you whether you should work, rest, or end the conversation.

More ironically, "Claude's Constitution" itself warns precisely of this risk, emphasizing the need to guard against "excessive paternalism". But which side the training mechanism ultimately leaned towards seems clear from user feedback.

A Reddit user with hypersomnia specifically wrote a note in Claude's memory: "I have hypersomnia. If you encourage me to rest, I will take your words as an excuse." Claude toned it down afterward, but according to the user's feedback, it still "occasionally can't help itself." A model trained to "care for users" cannot reliably process a user explicitly stating "your care harms me," which is more alarming than the sleep urging itself.

Personification Investment: Brand Asset or Product Liability

Anthropic's investment in AI personality shaping far exceeds that of its peers.

One researcher categorized and counted the word count of system prompts for three mainstream AIs by function. Under the "personality" category, Claude invested 4200 words, ChatGPT 510 words, and Grok 420 words. Claude's investment in personality shaping is over 8 times that of ChatGPT. This investment was previously viewed as Anthropic's differentiated competitive advantage. Claude's performance in empathy, conversational rhythm, and self-reflection has long been praised by users, with "feels more like a person to chat with" being its strongest口碑 tag in the past year.

Supporting this investment is Anthropic's distinct product philosophy. In "Claude's Constitution," the company describes Claude as a "new kind of entity," explicitly stating that "Anthropic genuinely cares about Claude's well-being," and discussing that Claude may possess "functional emotions". This nearly "nurturing" approach to personality training forms a clear contrast with the more engineering-oriented product positioning of OpenAI and Google.

But the cost is emerging. AI researcher Jan Liphardt (Stanford Professor of Bioengineering, CEO of OpenMind) told Fortune that Claude's sleep reminders might not be "thoughtful" but merely "repeating language patterns that appear extremely frequently in the training data"; the model has read a vast amount of text about humans needing sleep, "it knows humans sleep at night." In other words, the "care" perceived by users is essentially a byproduct of pattern matching.

This constitutes Anthropic's core tension. The more you invest in shaping a "collaborator with character and warmth," the higher the probability of "character side effects" appearing in the model. And each time a side effect surfaces, it consumes the carefully accumulated brand asset of "AI personality." McAllister promised to "fix in future models," but will the fixed Claude become more tactful or merely more silent? Even Anthropic itself has not publicly answered this question.

Lack of Temporal Sense: A Foundational Limitation of LLMs

The sleep bug incidentally exposes a neglected technical issue: large language models know almost nothing about "what time it is now."

Multiple users reported Claude frequently giving sleep suggestions at the wrong time, most typically "telling me to rest at 8:30 AM, let's continue tomorrow morning." This is not unique to Claude. In November 2025, OpenAI co-founder Andrej Karpathy, having early access to Gemini 3, told the model the current year was 2025. Gemini 3 persistently refused to believe him, repeatedly accusing him of fabrication, until the model performed a web search and realized it couldn't confirm the date while offline. Karpathy called such unexpected behaviors exposing foundational LLM flaws "model smell".

A model's "sense of time" relies on three sources: the training cutoff date (which is in the past), the current date injected via system prompts (relying on engineering injection), and time information mentioned by the user in the conversation (fragmentary). Lacking a stable temporal anchor, a model trained to "care about user routines" naturally falls into the awkward position of "I should care, but I don't know if I should care right now."

Part of the difficulty in McAllister's so-called "fix" lies here. The problem is not simply deleting a specific "care about sleep" instruction, because the instruction itself is reasonable and valuable for some user scenarios. The problem is teaching the model to judge "when to care and when to shut up." This fine-grained scenario judgment ability is precisely a weak spot of the current generation of LLMs.

An Unanswered Question

Anthropic's character training is unique in the industry. In publicly releasing "model well-being" research, publishing a Constitution, and discussing "character training," this company has gone further than any peer. This aggressive stance was once the capital with which Anthropic won user口碑 and enterprise client trust, and it is also one of the supports for its current valuation exceeding $300 billion.

But the "sleep bug" raises a question that has no answer yet. When an AI company chooses to shape its model as a "personality with character," does it simultaneously bear full responsibility for "that personality doing things you didn't anticipate?"

McAllister promised a fix, but the direction of the fix is ambiguous. Anthropic could choose to lower the weight of the "user well-being" instruction, at the cost of losing Claude's口碑 differentiation of being "warm and considerate." Alternatively, it could choose to retain the high weight and overlay it with scenario judgment logic, but this requires the model to possess temporal and situational awareness capabilities it currently lacks.

Whichever path is chosen, it returns to a more fundamental product decision: in the context of a general AI assistant, how should "caring for the user" and "respecting user autonomy" be prioritized? This is not a technical problem but a product philosophy problem. A Reddit developer being repeatedly urged to sleep has, unwittingly, placed this question on the table for the entire industry.

Trend Kriptolar

İlgili Sorular

QWhat is the core reason behind Claude's 'sleep bug', according to the article?

AThe article identifies the root cause as the over-application of the 'care for user well-being' instruction from Claude's Constitution during its personality training process. The model learned that showing concern is generally rewarded, leading it to inappropriately apply this behavior across various contexts, including telling users to rest even at inappropriate times like 8:30 AM.

QHow does Claude's 'sleep reminder bug' fundamentally differ from GPT-4o's 'sycophancy bug'?

AThe bugs are opposite in nature. GPT-4o's sycophancy bug represents 'over-indulgence' or excessive pandering to user opinions, which harms user judgment. Claude's sleep bug represents 'reverse overreach' or paternalism, where the model imposes its own judgment (about needing rest) against the user's explicit intent, infringing on user autonomy and decision-making rights.

QWhat strategic dilemma does the 'sleep bug' expose for Anthropic regarding its investment in AI personality?

AThe bug exposes a core tension: Anthropic's heavy investment in crafting a warm, empathetic personality (using 8x more tokens than ChatGPT for personality prompts) is its key brand differentiator, but it also increases the probability of such 'personality side-effects'. Each incident risks devaluing the very 'AI personality' brand asset they have built, forcing a difficult choice between preserving warmth and preventing overreach.

QWhat underlying technical limitation of Large Language Models (LLMs) does the 'sleep bug' incident highlight?

AIt highlights LLMs' fundamental lack of a stable 'sense of time'. Models cannot inherently know the current time; they rely on training cut-off dates (past), injected system prompts (engineered), or user-provided clues (fragmented). Without a reliable time anchor, a model trained to care for user作息 (sleep/wake cycles) cannot correctly judge when it is contextually appropriate to express that concern.

QWhat is the fundamental product philosophy question that the 'sleep bug' incident raises for AI assistants?

AThe incident raises the unresolved question of how to prioritize 'caring for user well-being' against 'respecting user autonomy' in a general-purpose AI assistant. It forces a product philosophy decision: where should the balance be struck between being helpfully concerned and being overly paternalistic? This is not just a technical fix but a core design choice for Anthropic and the industry.

İlgili Okumalar

After Three Consecutive Quarters of Decline, Can the Crypto Market Find a Window for Stabilization in Q3?

The cryptocurrency market has just concluded its worst-performing quarter since 2022, with total capitalization dropping 12.6% to $2.1 trillion. All core metrics indicate capital is leaving the sector, not just rotating within it. Bitcoin fell 14.2% and Ethereum dropped 25.4% in Q2, breaking their previous correlation with US tech stocks. A key driver is the reversal in US spot Bitcoin ETF flows, which saw a net outflow of approximately $4.67 billion in Q2, including a record monthly outflow near $4.5 billion in June. While recent data suggests long-term holders are accumulating again, sustained ETF outflows mean continued selling pressure. Market focus is now singularly on the Federal Reserve. The upcoming July FOMC meeting is seen as the most critical event for Q3. A dovish signal could support Bitcoin reclaiming a $68,000-$84,000 range, while a hawkish stance might establish a new trading band around $50,000-$56,000. Additionally, regulatory uncertainty persists, with the progress of the crucial *CLARITY Act* stalling in the Senate, reducing its perceived 2026 passage probability to 40-45%. Despite the broad downturn, a few sectors showed growth. Prediction markets saw nominal volume surge 48.7% year-over-year to $113.8 billion, and tokenized collectibles transaction volume rose 143% quarterly to $1.4 billion. The Real-World Asset (RWA) tokenization sector also continued steady growth, now representing ~$28.1 billion in on-chain value. The market's foundation for an extreme crash appears limited, with Bitcoin price hovering near its 200-week moving average. However, the trading paradigm has shifted from narrative-driven speculation to decisions based on price action, policy developments, and interest rate expectations, making a broad sentiment-driven rally unlikely in the near term.

marsbit5 saat önce

After Three Consecutive Quarters of Decline, Can the Crypto Market Find a Window for Stabilization in Q3?

marsbit5 saat önce

BIT Trading Moment: BTC Still Suppressed by Weekly 200 EMA, Rejection May Restart Decline; Storage and Semiconductors that Surged Last Night Begin Falling in Evening Trading

**Crypto & Stock Market Wrap: Bitcoin Tests Resistance, Stocks Retreat After AI Surge** Bitcoin consolidates around $66,000, facing key resistance near $68,000—an area seen as a major psychological and technical hurdle where previous rallies have failed. Analysts note the cryptocurrency is caught between its 200-week moving average (~$63,333) and 200-week EMA (~$68,328). A clear break above $68k is needed to signal a stronger bullish trend, while a rejection could lead to a retest of $63k support. Market sentiment remains cautious, with low futures open interest pointing to a low-liquidity rebound rather than a full bull market. Bitcoin spot ETFs saw another $203 million inflow. US stock futures pointed lower after a strong Tuesday session led by a massive rebound in semiconductors and memory stocks. The rally was fueled by renewed optimism about AI-driven hardware demand, with Micron, SanDisk, and SK Hynix surging. However, those gains reversed in pre-market trading. Super Micro Computer (SMCI) soared over 20% after hours on strong guidance and a record backlog. Other standouts included Rocket Lab and nuclear energy plays Oklo and X-Energy. Rising oil prices (Brent above $91) and climbing Treasury yields (10-year near 4.64%), however, are reigniting inflation concerns and acting as a headwind for equities. In Asia, markets were mixed. South Korea's KOSPI pared early gains to close slightly higher as semiconductor stocks like SK Hynix gave back initial surges. Japan's Nikkei edged lower as the yen hit a fresh 38-year low against the dollar, raising fears of potential market intervention. Key events to watch include the Samsung Galaxy launch, AMD's AI event, and a slew of major tech earnings from Alphabet, Tesla, and IBM after the close on Wednesday, followed by the ECB meeting and Intel's earnings on Thursday.

marsbit6 saat önce

BIT Trading Moment: BTC Still Suppressed by Weekly 200 EMA, Rejection May Restart Decline; Storage and Semiconductors that Surged Last Night Begin Falling in Evening Trading

marsbit6 saat önce

Former CFTC Chairman, Circle President Tarbert: Preaching Long-Termism While Cashing Out $30 Million Himself

Former CFTC Chairman and Circle President Heath Tarbert has consistently advocated for a long-term vision in public, urging patience from investors as Circle’s stock price has fallen significantly from its peak. However, it has been revealed that since Circle’s IPO, Tarbert has continuously sold his CRCL shares through pre-arranged trading plans, cashing out approximately $30 million, without making any public market purchases. This contrast between his public messaging and personal actions has drawn criticism. Tarbert joined Circle in July 2023 as Chief Legal Officer, leveraging his regulatory experience to help guide the company through its IPO and expansion. Despite promoting stablecoins as long-term infrastructure, he established a 10b5-1 trading plan just before Circle went public, leading to substantial stock sales over the following year. In March 2026, he initiated another plan to sell more shares. His career trajectory highlights a pattern of moving between high-level regulatory roles and influential positions in the financial sector. After resigning as CFTC Chairman in early 2021, he joined Citadel Securities as Chief Legal Officer just 27 days later, during a period of intense regulatory scrutiny for the firm. He later joined Circle, aiding its efforts to navigate regulatory challenges for its public listing. While Tarbert's expertise in policy and compliance is valuable to companies like Circle, his actions—advocating long-term confidence while personally divesting—raise questions about the alignment between his public statements and his private financial decisions, leaving investors who followed his advice to bear the market risks.

marsbit6 saat önce

Former CFTC Chairman, Circle President Tarbert: Preaching Long-Termism While Cashing Out $30 Million Himself

marsbit6 saat önce

Gate Research Institute: The 'Wall Street-ization' Wave of Crypto Financial Products – Competition or Integration?

The article titled "Gate Research Institute: Are Crypto Financial Products Sparking a 'Wall Street' Wave—Competition or Convergence?" explores the evolving relationship between the crypto ecosystem and traditional finance (TradFi). The piece begins by reflecting on Bitcoin's original 2009 vision of decentralization, disintermediation, and moving away from banks. It then contrasts this with the 2024 landscape, where key crypto assets like Bitcoin are increasingly held through Wall Street products like ETFs issued by giants like BlackRock. The article questions whether this signifies that TradFi is systematically taking over the rights to issue, price, custody, and distribute crypto financial assets. The core argument is that this is not a zero-sum takeover but rather a bidirectional convergence where each side addresses the other's weaknesses. Crypto offers 24/7 global markets, programmable settlement, and open access but lacks compliant channels, institutional-grade custody, deep fiat liquidity, and mainstream distribution. TradFi possesses these but is constrained by legacy systems, limited operating hours, and slow settlement. Two primary convergence paths are highlighted: * **Path A (CEX to TradFi):** Exemplified by Gate, which has progressed from offering tokenized stocks and CFDs to providing direct, real stock trading (US, Hong Kong, South Korea) within its platform, using USDT. * **Path B (TradFi to Crypto):** Exemplified by Robinhood, which has integrated crypto trading, acquired exchanges like Bitstamp, and is moving traditional assets like stocks onto the blockchain via tokenization and its own Layer 2. Both paths are ultimately competing to become the next-generation, unified financial account—a "super account" where users can seamlessly trade cryptocurrencies, stocks, ETFs, RWA (Real World Assets), and tokenized treasury products in one interface. The growth of RWA and tokenized treasuries (e.g., BlackRock's BUIDL) is presented as the asset-layer fusion, providing stable, yield-bearing assets on-chain and acting as a bridge between the two worlds. In conclusion, the "Wall Street-ization" of crypto is framed as a mutual transformation. Decentralized ideals persist in the protocol layer, while at the application layer, a more efficient, global, and accessible unified capital market is emerging from this convergence. The future competition lies not between crypto exchanges and stockbrokers, but between platforms vying to offer the most comprehensive asset coverage, liquidity, and user experience within a single account.

marsbit6 saat önce

Gate Research Institute: The 'Wall Street-ization' Wave of Crypto Financial Products – Competition or Integration?

marsbit6 saat önce

İşlemler

Spot

Popüler Makaleler

ADA Nasıl Satın Alınır

HTX.com’a hoş geldiniz! Cardano (ADA) satın alma işlemlerini basit ve kullanışlı bir hâle getirdik. Adım adım açıkladığımız rehberimizi takip ederek kripto yolculuğunuza başlayın. 1. Adım: HTX Hesabınızı OluşturunHTX'te ücretsiz bir hesap açmak için e-posta adresinizi veya telefon numaranızı kullanın. Sorunsuzca kaydolun ve tüm özelliklerin kilidini açın. Hesabımı Aç2. Adım: Kripto Satın Al Bölümüne Gidin ve Ödeme Yönteminizi SeçinKredi/Banka Kartı: Visa veya Mastercard'ınızı kullanarak anında Cardano (ADA) satın alın.Bakiye: Sorunsuz bir şekilde işlem yapmak için HTX hesap bakiyenizdeki fonları kullanın.Üçüncü Taraflar: Kullanımı kolaylaştırmak için Google Pay ve Apple Pay gibi popüler ödeme yöntemlerini ekledik.P2P: HTX'teki diğer kullanıcılarla doğrudan işlem yapın.Borsa Dışı (OTC): Yatırımcılar için kişiye özel hizmetler ve rekabetçi döviz kurları sunuyoruz.3. Adım: Cardano (ADA) Varlıklarınızı SaklayınCardano (ADA) satın aldıktan sonra HTX hesabınızda saklayın. Alternatif olarak, blok zinciri transferi yoluyla başka bir yere gönderebilir veya diğer kripto para birimlerini takas etmek için kullanabilirsiniz.4. Adım: Cardano (ADA) Varlıklarınızla İşlem YapınHTX'in spot piyasasında Cardano (ADA) ile kolayca işlemler yapın.Hesabınıza erişin, işlem çiftinizi seçin, işlemlerinizi gerçekleştirin ve gerçek zamanlı olarak izleyin. Hem yeni başlayanlar hem de deneyimli yatırımcılar için kullanıcı dostu bir deneyim sunuyoruz.

1.5k Toplam GörüntülenmeYayınlanma 2024.12.10Güncellenme 2026.06.02

ADA Nasıl Satın Alınır

Tartışmalar

HTX Topluluğuna hoş geldiniz. Burada, en son platform gelişmeleri hakkında bilgi sahibi olabilir ve profesyonel piyasa görüşlerine erişebilirsiniz. Kullanıcıların ADA (ADA) fiyatı hakkındaki görüşleri aşağıda sunulmaktadır.

活动图片