Excess-vocabulary cluster (delve, underscore, intricate)
- Id
- delve-excess-vocabulary
- Status
- Fading
- Severity
- medium
- Detection
- statistical
- Evidence grade
- peer-reviewed
- Languages
- en
- Added
- 2026-08-14
- Updated
- 2026-08-17
Signal is weakening, usually because model vendors trained the habit out.
What it is
A set of ordinary English style words whose frequency jumped after late 2022. Kobak and colleagues measured the jump across more than 15 million PubMed abstracts, with delves at about 25 times its extrapolated pre-2022 baseline. The words are not new. What changed is how often several of them arrive in the same short passage, which is the effect Kousha and Thelwall tracked across six scholarly databases. The detection here is a density count, not a word ban: distinct markers drawn from the two published lists, counted per 1000 words. The threshold of 3 is a working default, not a published cutoff. It is set there because the published lists contain common academic words with real human base rates, so one hit carries almost no information and Wikipedia's own discipline rule is that co-occurrence is the signal. Recalibrate it against a measured corpus before treating it as a number rather than a starting point.
Why it reads as machine-made
One marker word means nothing. The cluster is what readers react to: several of these words in a paragraph with no proper noun or number between them. The studies behind the list measured populations, not documents. A sentence containing delve tells you about the corpus it came from and almost nothing about who typed it.
Specimens
This piece delves into the intricate landscape of remote onboarding, underscoring the pivotal part a robust welcome sequence plays in long-term retention. The interplay between structure and warmth is what most teams eventually garner from the exercise. Remote work remains a vibrant realm, and the organisations that treat it as one tend to bolster retention without ever quite saying how they did it. Practitioners would do well to interrogate their assumptions before wholesale adoption, since what proves efficacious in one setting may prove considerably less so in another.
Our welcome sequence is four emails over nine days. The third one asks the reader to name the job they hired the product for. Those answers predict cancellations better than anything in our analytics.
The repair is not synonym swapping. The specimen states nothing about the sequence at any point. The after text says how long it runs and what it is used for.
How it is detected
- Metric
- distinct-excess-vocabulary-markers-per-1000-words
- Threshold
- 3
- Direction
- above
- Threshold basis
- Working default, not a published cutoff. Kobak and colleagues measure excess vocabulary across 15 million abstracts at corpus scale and publish no per-document threshold, because the method is population-level by design. Three distinct markers per 1,000 words is a FeedSquad review trigger chosen so that one ordinary use of one word cannot fire it. It flags text for a human to read. It decides nothing.
Who writes this way legitimately
Biomedical and STEM academics used these words before 2022. Kobak's method measures excess over an extrapolated baseline, which means the baseline was real and nonzero. Nigerian English speakers publicly objected in April 2024 that delve into is ordinary business register for them, and no corpus study settling that question has been published, so the dispute stands open. Podcast speakers now produce the cluster in unscripted speech, per Yakura. Non-native English writers are the group with most to lose: Liang and colleagues measured a 61.3% average false-positive rate across seven detectors on human-written TOEFL essays, and the prompt that cleared the false accusation was one that made the vocabulary fancier.
Model attribution
Corpus-level LLM-era signal, not a family marker. Neither Kobak nor Liang assigns any word to a model family. Wikipedia's era buckets suggest the vocabulary drifted between model generations, and that periodisation is editor observation rather than measurement.
Platform notes
- wikipedia
- Wikipedia keeps a per-word sourced list and states the rule literally: a word being overused by AI does not imply its synonyms are. Editors must corroborate a word from a non-pop-science source before adding it.
- No LinkedIn policy names any word. The May 2026 announcement targets posts with no unique perspective and reduces distribution outside a person's network rather than removing the post.
Status history
| Date | Status | Rationale |
|---|---|---|
| 2026-08-14 | Fading | The Washington Post measured delve in roughly 1 in 1,000 publicly shared ChatGPT messages by July 2025, well below its 2023 peak. Yakura and colleagues found the same words rising in unscripted human speech after ChatGPT's release, with a preregistered experiment showing adoption after brief exposure. The cluster still measures something at corpus scale. It no longer separates one author from another. |
Sources
- 01
- 02Same study, arXiv HTML v1 (excess frequency ratios and gaps)primary-docaccessed 2026-08-14
- 03Kousha and Thelwall: How much are LLMs changing the language of academic papers after ChatGPT?primary-docaccessed 2026-08-14
- 04Yakura et al.: Empirical evidence of Large Language Model's influence on human spoken communicationprimary-docaccessed 2026-08-14
- 05
- 06
- 07Vanguard Nigeria: Nigerians tackle American author for claiming delve is only used by ChatGPTpressaccessed 2026-08-14
- 08Wikipedia: Signs of AI writingcommunityaccessed 2026-08-14
- 09Juzek and Ward: Why Does ChatGPT Delve So Much? (COLING 2025)peer-reviewedaccessed 2026-08-14
- 10Geng and Trotta: coevolution of human and LLM writing (Findings of ACL 2025)peer-reviewedaccessed 2026-08-14
- 11LexA-Index, CC0 per-language overuse datasetprimary-docaccessed 2026-08-14
CC BY 4.0 / The AI Tells Index, feedsquad.com/ai-tells