One-Third of arXiv 'Contaminated', 65% of CS Papers Smell of AI, Only 0.7% in Math

marsbitPublished on 2026-07-27Last updated on 2026-07-27

Abstract

Approximately one-third of recently posted arXiv papers show signs of significant AI-generated text, according to a new study. An analysis of 12,750 papers from January 2023 to July 2026 across ten disciplines found a sharp increase in AI text markers following ChatGPT's release, with the overall detection rate reaching 32% in the latest quarter and peaking near 39% in early 2026. The rate varies drastically by field. Computer Science papers lead at 65%, followed by Quantitative Biology (56.3%) and Electrical Engineering (51.3%). Mathematics, however, has the lowest detection rate at just 0.7%. The study's authors note this could be due to mathematicians using AI less or because the detector struggles with the high volume of formulas and symbolic notation in math papers, leaving the true cause unclear. The research highlights that the detector identifies a statistical "AI style" in the text rather than proving full AI authorship. It cannot distinguish between light AI-assisted editing and fully AI-generated content. Furthermore, the detector can produce false positives, as some pre-ChatGPT academic writing also exhibits patterns now flagged as "AI-like." The growing use of AI, particularly in highly competitive fields, is creating a cycle where researchers may feel pressured to adopt AI tools to keep pace. The findings raise questions about the changing nature of academic writing and the emergence of a new "AI style" that is increasingly difficult to distinguish from human...

A third of arXiv papers are actually written by AI?

A recent unslop study has caused a stir overseas.

Over the past year, 65% of new papers in computer science have been flagged by its detector.

Netizens have even coined a term for it – 'The Great Slurry Era'.

Yet, for the same period, only 0.7% of math papers were flagged.

ChatGPT Appears, Curve Takes Off

In this study, the team scanned 12,750 arXiv papers spanning from January 2023 to July 2026.

The target was ten disciplines, with about 25 full-text papers sampled per field per month.

The control group was set from 2021 to 2022, the pure human era before ChatGPT's birth.

The results show that from 2021-2022, the flagging rate was stable at 0.4%; but with ChatGPT's release, the curve shot up within months.

In the most recent complete quarter, the flagging rate reached 32%, and in early 2026, the peak approached 39%.

By discipline, computer science fared the worst, with a flagging rate of 65%.

Before ChatGPT, its baseline was only 0.2%. In just three years, the number soared over 300-fold.

Following closely were quantitative biology at 56.3%, electrical engineering at 51.3%, economics & finance at 47%; then applied physics at 34%, statistics at 31.3%, condensed matter physics at 24%, high-energy physics at 14%, and astrophysics at 10.7%.

The bottom of the list was mathematics: 0.7%.

And this is just the lower bound. If AI text is hidden well enough, the detector can't identify it.

The same detector, the same paper database, a gap nearing 100-fold.

Are mathematicians collectively sticking to pure manual writing? Or is this machine simply blind in the face of mathematical formulas?

The Bottomed-Out 0.7%: Machine Blindness

In fact, the answer has been written into the unslop report itself.

Think, what does a standard mathematics paper look like? Screens full of symbols, formulas, theorems, and proofs. The genuine prose sentences written in English are pitifully few.

This detector only recognizes text. It focuses on sentence structure, word usage habits, paragraph organization.

And the few remaining lines of text in math papers are all formulaic expressions like 'Let X be such and such' or 'By Lemma 3, we can derive', which basically aren't in the same realm as the scientific English the detector was trained on.

So the team itself admits the study currently has several limitations:

The 0.7% for math doesn't prove much yet. It might be that mathematicians really don't use AI much, or the detector simply can't read math papers, but this data can't distinguish which.

The control group is also relatively small. Only 200 pre-ChatGPT papers per discipline, with a 0.4% false positive rate, means only about 8 flagged papers out of 2000 total. This level can only be considered a rough estimate.

Then there's the aforementioned issue of incomplete detection. So, these numbers are just lower bounds; the real proportions are certainly higher.

The More Competitive the Field, the More Reliance on AI

So, who is using AI to write papers?

The answer lies in another two-year Stanford study.

In March 2024, Stanford's Weixin Liang team first measured peer reviews: in ICLR 2024 review comments, about 10.6% of sentences were heavily modified by LLMs. The reviewers used it themselves before even finishing evaluating the papers.

In August 2025, the same team expanded the sample to 1.12 million papers and directly published in Nature Human Behaviour.

The research showed that as of September 2024, the proportion of LLM-modified text in CS paper abstracts had reached up to 22.5%. And the heaviest users were researchers publishing preprints most frequently, operating in the most competitive fields.

The more competitive the field, the heavier the use.

Here, AI writing has become an arms race: peers are using AI to speed up; those writing by hand alone fall a step behind.

No wonder a widely circulated related repost on X had the first line: 'Prompt: Write a paper publishable on arXiv.'

It Sniffs the 'Smell', Not the AI

But don't rush to convict based on that 65%.

The team also admits in the report that the detector cannot distinguish between 'AI polished the grammar' and 'the whole thing was machine-made'. It only detects the concentration of machine-like scent in the text.

So the correct understanding of 65% is '65% of papers smell like AI', not '65% of papers are written by AI'.

The trouble is, humans themselves can acquire this machine scent.

A skeptical researcher fed their own old papers from years ago into a detection tool, and 27% to 74% of the content was flagged red.

Remember, in that era, ChatGPT wasn't even a glimmer.

The reason isn't hard to find.

Open any academic journal today, and pages are filled with large language models, benchmark tests, state-of-the-art. The more standardized the phrasing, the more structured the composition, the more likely it is to hit the machine-like features recognized by the detector.

Besides, LLMs were trained on such papers. It's not that scholars write like AI; it's that AI was born writing like scholars.

When All Text Becomes Suspect

Actually, in recent years, we've all been doing the same thing as the detector.

Reading a passage that's perfectly balanced, neatly phrased, occasionally popping with 'not only... but also...', our hearts skip a beat: 'This must be AI-written, right?'

Unslop's detector essentially turned this mystical sense of smell into an instrument that mass-produces numbers.

After all, once suspicion enters the mind, it's hard to remove.

In the past, 'writing it down' was itself an endorsement. Someone willing to spend hours organizing text at least showed seriousness.

Now, no matter how beautifully written, it might only earn a 'AI, right?'.

Some have even started deliberately avoiding words they've used all their lives, simply because they 'sound too AI'.

Now, 'AI scent' is becoming a new type of original sin sweeping through the world of writing.

The ironic part is, to this day, we still haven't precisely measured what exactly constitutes this 'AI scent'.

This article is from WeChat public account 'Xinzhiyuan', author: ASI Revelation

Trending Cryptos

Related Questions

QAccording to the article, what percentage of recent arXiv papers in computer science were flagged as having 'AI flavor' by the detection tool?

A65% of recent arXiv papers in computer science were flagged.

QWhy might the detection tool show a very low flag rate (0.7%) for mathematics papers compared to other fields?

AThe tool analyzes text structure and language, but mathematics papers are dominated by symbols, formulas, and theorems, with very little prose. The limited text often uses formal, standardized phrasing that the tool may not recognize as typical scientific English, potentially making it 'blind' to AI use in this context.

QWhat does the Stanford study mentioned in the article suggest about the correlation between AI use and research field competitiveness?

AThe Stanford study suggests that the most competitive research fields, where researchers publish preprints most frequently, show the highest levels of LLM-assisted writing. It has become an 'arms race' where using AI provides a speed advantage.

QWhat is a key limitation of the AI-text detection tool discussed in the article?

AA key limitation is that the tool detects a general 'AI flavor' or style in the text but cannot distinguish between partial AI use (like grammar polishing) and fully AI-generated content. It also has high false positive rates, as it sometimes flags human-written academic prose.

QWhat broader societal concern regarding writing does the article raise in its conclusion?

AThe article raises the concern that 'AI flavor' is becoming a new kind of 'original sin' in writing. Well-written, clear, and structured text is now often met with suspicion of being AI-generated, which undermines the trust and value traditionally placed on carefully crafted human writing.

Related Reads

Must-Watch Events Next Week|CLARITY Act Could Face Senate Vote; SpaceX, Circle to Report Earnings (8.3-8.9)

**Summary: Key Events and Developments to Watch (August 3-9)** The upcoming week is marked by significant financial disclosures, key legislative deadlines, and notable product updates. **Major Financial Events:** Several companies are scheduled to release their Q2 2026 earnings. American Bitcoin (ABTC) will report on August 3, followed by SpaceX and Hut 8 Mining Corp. on August 4, and Circle on August 5. Notably, a significant portion of SpaceX shares (up to 12% of total shares) will be unlocked on August 6 following their earnings release. **Key Legislative Deadline:** The U.S. Senate faces an August 7 deadline to secure 60 votes for the CLARITY Act, a bipartisan bill aiming to establish a federal regulatory framework for cryptocurrencies. The Senate may hold a full vote on the bill during the week. **Economic Data:** The U.S. July Non-Farm Payrolls report will be released on August 7, providing crucial labor market data. **Technology & Product Updates:** * **Shutdowns:** DeFi portfolio tracker Zapper and wallet app Ctrl Wallet will cease operations on August 3. * **Upgrades:** LayerZero will deprecate its v1 relayers on August 3. XRP Ledger's new version 3.3.0, featuring five new functions, is expected next week. * **AI:** Elon Musk announced that the advanced Grok 4.6 AI model is set for release around August 7. * **Bitcoin:** The BIP-110 forced signaling for a potential Bitcoin network change is scheduled to begin around August 8. **Other Notable Events:** Chinese robotics firm Unitree Tech has set its preliminary price inquiry for its IPO for August 5. South Korean exchange Upbit will delist AQT and AERGO tokens on August 3.

marsbit45m ago

Must-Watch Events Next Week|CLARITY Act Could Face Senate Vote; SpaceX, Circle to Report Earnings (8.3-8.9)

marsbit45m ago

Stocks Are Plummeting More Sharply Than Cryptocurrencies. Where Has the Money Gone?

Stock Markets Plunge Deeper Than Cryptocurrencies: Where Did the Money Go? In late July, Seoul's Kospi index triggered circuit breakers for two consecutive days, plummeting over 40% from its June high. The collapse was led by heavyweight stocks like SK Hynix, whose record profits still disappointed investors, and devastating leveraged ETFs, with one major product losing over 83% of its value. This signaled a global, forced deleveraging targeting the most crowded trades. Interestingly, while stocks exhibited extreme volatility akin to crypto markets, Bitcoin rose nearly 15% in July after a prior steep drop. Analysis shows the money fleeing equities did not flow into Bitcoin. Instead, Bitcoin had already absorbed its sell-off in May-June, when U.S. spot Bitcoin ETFs saw historic outflows. The true safe-haven beneficiary was gold, whose price rose over 20% year-on-year, highlighting a decoupling between Bitcoin and gold as "digital gold." The sell-off was a targeted unwinding of leveraged positions in tech and semiconductors, accelerated by broker-dealer risk management and shifts in the AI narrative, including new competition from Chinese memory chipmakers. The retreat path was clear: from high-valuation tech stocks to cash and U.S. Treasuries, then to gold. For Bitcoin to attract sustained institutional inflows, conditions like eased global liquidity pressure, a "soft-landing" Fed rate cut, and U.S. regulatory clarity via legislation like the stalled CLARITY Act are needed. Currently, Bitcoin is not a safe haven but an already-cleared asset. Its low correlation with tech stocks, however, makes it a potential diversification play for institutional portfolios once the storm passes. The money isn't here yet, but the positioning is underway.

marsbit45m ago

Stocks Are Plummeting More Sharply Than Cryptocurrencies. Where Has the Money Gone?

marsbit45m ago

In Conversation with Ray Dalio: We Are Currently in an AI Bubble, with 1% of My Portfolio in Bitcoin

Ray Dalio, founder of Bridgewater Associates, warns in an interview that the current AI boom shows classic bubble characteristics, which could lead to significant economic downturns as seen in past cycles like 1929 or 2000. He explains that speculative enthusiasm, fueled by debt and overvaluation, often precedes a crash when rising rates or taxation force asset sales, causing widespread losses and recession. Dalio also outlines his "Big Cycle" theory, describing an approximate 80-year pattern where widening wealth gaps, massive government deficits, and shifting geopolitical power (like China's rise) create internal conflict and global instability. He emphasizes that we are in a late-cycle, transitional phase where traditional powers like the US and UK face decline. For personal wealth protection, Dalio advises diversification beyond cash into assets like stocks, bonds, real estate, and particularly gold, which he prefers over Bitcoin. While he holds about 1% of his portfolio in Bitcoin as a non-printable hard asset, he views gold as more secure from technological or governmental threats. Regarding AI's impact, Dalio believes it will disproportionately benefit capital owners, worsening inequality by replacing both physical and cognitive labor. He suggests that human intuition and emotional intelligence, combined with AI, will be key for future workers. On taxation, Dalio argues that wealth taxes are impractical and risk triggering asset sell-offs, reducing productive investment. He points to the UK as a cautionary example of debt, low productivity, and political strife. Geopolitically, Dalio foresees a more regionalized world, with the US showing weakness in prolonged conflicts like with Iran, akin to past imperial declines. The ideal outcome, he suggests, is coexisting powerful blocs (e.g., Americas, China-Asia Pacific) without major war.

marsbit4h ago

In Conversation with Ray Dalio: We Are Currently in an AI Bubble, with 1% of My Portfolio in Bitcoin

marsbit4h ago

Daily 7.2 Trillion KRW: Foreign Capital's Record Net Buying on Friday! Wall Street Says Headwinds for Korean Stock Fund Flows Have Subsided

South Korean stock market sees a dramatic shift in fund flows. On July 31, foreign investors made a record net purchase of approximately KRW 7.2 trillion in KOSPI stocks, marking a fundamental reversal from the persistent large-scale net outflows seen in previous months. This contributed to a significant narrowing of foreign net selling in July to KRW 9.8 trillion, down sharply from KRW 48.4 trillion in June and KRW 44.5 trillion in May. Simultaneously, domestic institutional pressure eased. South Korean pension funds and asset managers turned to a net buying position in July, purchasing KRW 1.0 trillion worth of KOSPI shares, contrasting with net sales in May and June. Market volatility is expected to be dampened by new financial regulations. Effective July 31, the Financial Services Commission tightened access for retail investors to single-stock leveraged ETFs by raising the minimum cash deposit requirement. Trading volumes for these products subsequently dropped to about 50% of their monthly average. Citigroup Research maintains its year-end KOSPI target of 10,000 points. The firm cites several supportive factors: the substantial easing of headwinds from capital outflows, a robust fundamental outlook for the semiconductor sector, historically low market valuations, strong economic fundamentals, and the potential for policy support from financial authorities if needed.

marsbit4h ago

Daily 7.2 Trillion KRW: Foreign Capital's Record Net Buying on Friday! Wall Street Says Headwinds for Korean Stock Fund Flows Have Subsided

marsbit4h ago

Trading

Spot

Hot Articles

Discussions

Welcome to the HTX Community. Here, you can stay informed about the latest platform developments and gain access to professional market insights. Users' opinions on the price of AI (AI) are presented below.

活动图片