Where It Started
In 2021, an anonymous post appeared on Agora Road's Macintosh Cafe making an unusual claim: most internet activity is bots, and real humans have long become a minority online. The author called this the "dead internet."
At the time, it sounded like conspiracy theory. Today, it sounds like a technical report.
What the Theory Claims
The dead internet theory makes three assertions:
- Most content is generated automatically — for SEO, traffic inflation, and trend manipulation
- Most interactions — likes, comments, shares — are produced by bots
- Platform algorithms deliberately amplify synthetic content because it's predictable and controllable
When the theory emerged, evidence was scarce. Now there's plenty.
What Changed With LLMs
Before 2022, generating convincing text required significant effort. GPT-3 existed but was expensive and complex to use. After ChatGPT, the barrier dropped to zero.
| Year | Estimated AI content share on the web |
|---|---|
| 2019 | < 1% |
| 2022 | ~3–5% |
| 2024 | 15–20% (Cloudflare estimates) |
| 2026 (forecast) | > 50% |
This isn't neutral statistics — it's a change in the nature of the medium.
Three Levels of the Problem
The Dead Internet Threat Structure
Level 1: Information Noise
Search engines began indexing billions of AI-generated pages created purely for traffic capture. The term «AI slop» entered common usage precisely for this — content that looks like information but isn't.
Level 2: Social Engineering
Bot farms on social networks can now hold coherent conversations, simulate emotional reactions, and create a false sense of consensus.1 Distinguishing a real person from a language model is becoming increasingly difficult.
Level 3: Training Data Poisoning
Most troubling: the next generation of language models trains on an internet that already contains content from previous generations. The circle closes.
"We risk creating an echo chamber where AI learns from AI, progressively drifting further from reality."
Why This Matters for Enterprise AI
If your system makes decisions based on data from open sources — news, industry publications, forums — you're already operating on a partially synthetic reality.
- Analytics systems produce a distorted picture of the market
- RAG pipelines built on external sources degrade over time
- Sentiment models react to opinions that don't exist
The solution isn't paranoia — it's information hygiene: prioritize internal data, verify sources, isolate training corpora.
Our Perspective
We build AI systems that operate on closed, verified enterprise data. Not because we fear the internet — but because proprietary data is always more accurate, current, and reliable than any open source.
The dead internet isn't a reason to panic. It's an argument that a company's own data is a strategic asset.2
Conclusion
The dead internet theory started as conspiracy. Today it describes a measurable technical reality. The question isn't whether this is happening — it's how you adapt your information systems to an environment where synthetic content has become the norm.
Footnotes
-
According to Stanford Internet Observatory research (2023), in several political campaigns on Twitter/X, up to 80% of "supportive" comments were generated by coordinated bot networks with LLM backends. ↩
-
Gartner forecasts that by 2027, companies with verified proprietary data corpora will demonstrate 35–40% higher accuracy in enterprise AI systems compared to companies relying on public sources. ↩


