Content pruning means removing, merging or hiding pages that do not serve your readers, so the rest of the site is easier for people and search engines to use. Done carefully, it cleans up thin, duplicate and outdated pages that compete with your best content. Done carelessly, it deletes pages that were quietly earning links, conversions or long-tail traffic. This guide covers when pruning makes sense, what Google itself says about deleting content, and a step-by-step process to do it safely.
What Google says about removing content
Google’s guidance is more cautious than most pruning advice. Its page on core updates says to consider first how you can improve content in meaningful ways, and calls deleting content a last resort, only for content that can’t be salvaged. It adds that if you are considering deleting whole sections, that is likely a sign those sections were created for search engines first, and in that case deleting the unhelpful content can help the good content on your site perform better.
Two practical consequences:
- Improve before you delete. A page with a real purpose but weak content is usually worth rewriting, merging or updating.
- Pruning pays off most where content was made for search engines, such as mass-produced near-duplicate pages, auto-generated tag archives or thin articles written only to target a keyword.
On crawling, Google’s crawl budget guide is aimed mainly at large sites with over a million pages, or over 10,000 pages that change daily. For a site of a few hundred pages, crawl budget is rarely the reason to prune; quality and clarity are.
When pruning makes sense
- Several pages answer the same question. Overlapping articles split clicks and links between them. See how to fix keyword cannibalization.
- Pages were produced for search rather than readers: near-identical location or product variations, tag and archive pages with one or two posts, thin articles with no original information.
- Content is outdated and no longer true: old product versions, expired offers, events long past, advice that has since changed.
- Large numbers of pages are not indexed. In Search Console, many URLs under “Crawled, currently not indexed” or “Discovered, currently not indexed” can point to pages Google does not consider worth indexing.
When it does not
- Pages that serve users regardless of search: pricing, contact, legal, help and support pages, sign-up and checkout flows. Low search traffic is not a reason to remove them.
- Seasonal content that is quiet for most of the year and important for a few weeks.
- New pages that have not had time to earn traffic.
- Pages that convert or support conversions even with few visits. Check GA4 before judging a page by clicks alone.
The four decisions for each page
- Keep and improve. The page has a clear purpose but weak or outdated content. Rewrite, expand, update facts and dates, improve internal links to it.
- Consolidate. Two or more pages cover the same topic. Merge the best parts into one page, then 301 redirect the others to it.
- Keep for users, hide from search. The page is useful on the site but not as a search result, such as internal search results or filter combinations. Use
noindex. - Remove. The page has no value to users and nothing worth merging. Return 404 or 410; Google’s crawl budget guide recommends a 404 or 410 status for permanently removed pages.
How to run a content audit
1. Build the inventory
Combine three lists: a crawl of the site, the URLs in your XML sitemaps, and the pages Search Console reports. Each catches pages the others miss, such as orphan pages that only exist in the sitemap.
2. Collect data for every URL
- Clicks and impressions over the last 12 months, from Search Console.
- Engagement and conversions, from GA4. See GSC vs GA4 for what each measures.
- Backlinks, from your link tool.
- Internal links pointing to the page, and the date it was last updated.
- Word count and the main topic or query it targets, to spot overlaps.
3. Classify
Set your own thresholds for your site’s size and traffic; there is no universal number. A cautious starting rule for the “remove” group is: no clicks in 12 months, no backlinks from relevant sites, no conversions, no purpose for users, and nothing worth merging into another page. Anything that fails only some of these tests belongs in “improve” or “consolidate”.
4. Decide and record
Write the decision and the target URL for each page in one sheet before you change anything. It is your plan, your redirect map and, if something goes wrong, your way back.
Tip: before removing any page, check its backlinks. A page with no traffic can still have a link from a relevant site. Redirect it to the closest equivalent page instead of letting the link point at a 404.
Redirects done right
- Redirect to the closest equivalent. A merged article goes to the article it was merged into; an old product to its replacement.
- Do not send everything to the homepage. Google warns that redirecting many old URLs to one irrelevant destination, such as the home page, can confuse users and might be treated as a soft 404. If there is no relevant page, a 404 or 410 is the honest answer.
- Update internal links to point straight at the final URL, so you do not build redirect chains.
- Remove pruned URLs from the sitemap, and keep redirected URLs out of it too.
- Keep redirects in place long term. Links and bookmarks to old URLs keep arriving for years.
What to expect, and how to measure it
Do not expect a fixed percentage. Google says that after improvements some changes can take effect in a few days, but it could take several months for its systems to confirm that a site as a whole is now producing helpful content. Measure rather than guess:
- Clicks and impressions of the pages you kept in Search Console, compared over the same period before and after.
- The Page indexing report: fewer excluded URLs and no new soft 404s from your redirects.
- Crawl stats, if your site is large enough for crawl budget to matter.
- Conversions in GA4, to make sure nothing that sold was removed.
Work in batches, starting with the clearest cases, note the date of each batch, and wait for data before the next one. Pruning too much at once is the most common way to make things worse, because you cannot tell which change caused what.
Common pruning mistakes
- Judging pages by clicks alone. A page with few clicks may rank for queries that convert, attract links, or answer a support question your customers need. Always look at links, conversions and purpose as well.
- Removing pages that have backlinks without redirecting them. The links then point at an error page and stop helping the site.
- Blocking pages in robots.txt and adding noindex at the same time. Google says that for noindex to work, the page must not be blocked by robots.txt; otherwise the crawler never sees the noindex rule and the page can still appear in results.
- Pruning everything in one go. If traffic drops afterwards, there is no way to tell which change caused it, and no clean way back.
- Forgetting the leftovers: internal links still pointing at removed pages, removed URLs still in the sitemap, and menus or related-post widgets that link to them.
- Merging without merging. Redirecting three articles to a fourth that does not actually include their useful content loses what made them worth visiting.
Content decay is not the same as pruning
A decaying page once performed well and is now losing traffic, usually because it is outdated or competitors published something better. It has proven value: update it, do not delete it. Pruning is for pages that never had a clear purpose or have lost it for good. Treat the two lists separately so you do not remove a page that only needed a refresh.
For the full technical side of a site review, see the technical SEO audit checklist, and for how internal links spread value across the pages you keep, internal linking architecture.