Thin Content vs. Duplicate Content: two different problems
The two terms get mixed up easily because they show up in the same audits and both get filed under "low-quality content." But they answer different questions. Thin content asks how much value a piece of text gives the person reading it: does it solve their question, does it tell them something new, does it justify the click. Duplicate content asks whether that text already exists somewhere else, identical or nearly so, regardless of whether it's good or bad.
A piece of text can be weak and completely original at the same time: a hundred-word product page, written from scratch, that only says "this bag is elegant and practical" without a single measurement, material, or fact that helps someone decide to buy it. Nobody else has that exact paragraph, so there's no duplicate content problem. But it doesn't deliver anything either, so it's still thin content. The reverse also happens: a well-researched article, with real data and a solid structure, that another site copied and republished unchanged on its own domain. That second site doesn't have a depth problem, it has an originality problem.
The two dimensions are independent of each other, which is why it helps to picture them as two crossing axes: how much value the content delivers, and whether it's exclusive to that URL or not.
| Unique content | Duplicated content | |
|---|---|---|
| Rich content | The ideal case: complete, original, well-structured information that doesn't exist with this exact wording anywhere else. | Genuinely good content republished unchanged on another domain, such as a press release syndicated across several outlets. |
| Thin content | Weak but with no duplicate content problem: a short, original page that still says nothing useful, even though nobody else has it. | Both problems at once: mass-produced product pages that also copy the manufacturer's description word for word. |
The box that causes the most headaches in practice is thin and duplicated at the same time. It's the typical pattern of a large e-commerce catalog: thousands of product pages carrying the same manufacturer description, with no added value of their own, also replicated across dozens of other stores selling the same product. A canonical tag alone doesn't fix that, because the underlying problem, the missing original content, is still there even after Google consolidates the signals into a single URL.
