What Index Bloat Is
Index bloat is the situation where Google has indexed far more of a site's URLs than that site needs. It is not a specific error you can point at on a page but an imbalance: the index fills with addresses that contribute nothing, and the content that matters gets diluted among them.
The number alone says nothing. A newspaper with two hundred thousand indexed URLs is fine; a shop with three hundred products and twenty thousand is not. What counts is not how many there are but what share of them anyone ever searches for.
And the opposite belief deserves clearing away at the outset, because it persists: more indexed pages improve nothing by themselves. Google does not hand out visibility by volume, and an indexed URL nobody visits is not an asset but spent crawling.
The name describes what happens well: something swells without gaining substance. There is no threshold at which it begins and no warning in any tool; you recognise it by comparing two figures hardly anyone puts side by side — the pages you believe you have and the URLs Google says it has.
It also deserves separating from a neighbouring problem it gets confused with. A page being indexed and receiving no visits can be perfectly normal: there is seasonal content, legal pages, product pages for rare items. Index bloat is not that; it is the case where such URLs run into the thousands and none of them was planned.
