SEO Content Audit: The Five Decisions Google Documents, and the Arithmetic That Says Where to Spend the Hours
· Royking Niba
An SEO content audit is the process of taking every content URL on a site and assigning each one a decision. There are only five decisions available: leave it alone, improve it, consolidate it into another page, keep it but remove it from search, or delete it. Most audits fail not because the inventory is wrong but because they treat all five as equally available and all URLs as equally worth the time. Google’s documentation is unusually direct about one of those decisions, and basic arithmetic settles the other half of the problem.
What Google actually says
From Google’s core updates guidance, checked on 2 October 2026, on what to do with content that is not performing:
Deleting content is a last resort, and only to be considered if you think the content can’t be salvaged.
Google Search Central, Google Search’s core updates and your website
That single sentence invalidates most of what gets sold as a content audit, because the deliverable is usually a delete list. The same page adds three more constraints worth holding onto. It says to "take a close look at your site as a whole, and try to be objective", which makes the audit a site-level exercise rather than a URL-level one. It says to "avoid doing ‘quick fix’ changes (like removing some page element because you heard it was bad for SEO)". And it sets the honest expectation on timing: "some changes can take effect in a few days, but it could take several months for our systems to learn and confirm that the site as a whole is now producing helpful, reliable, people-first content."
It also carries the sentence every audit proposal should quote and almost none do: "There’s no guarantee that changes you make to your website will result in noticeable impact in search results."
For judging an individual page, Google’s helpful content guidance supplies the test. The questions it asks are about originality and usefulness rather than length: "Does the content provide original information, reporting, research, or analysis?" and "Is this the sort of page you’d want to bookmark, share with a friend, or recommend?" It asks whether the page avoids "simply copying or rewriting" other sources without "substantial additional value", and whether it shows "clear sourcing, evidence of the expertise involved". On the Who, How and Why framing it is blunt about the third: "The ‘why’ should be that you’re creating content primarily to help people."
One more question from that page kills a tactic that still appears in audit templates: "Are you changing the date of pages to make them seem fresh when the content has not substantially changed?" A refresh pass that only touches the publish date is not an improvement, and Google names it as a problem signal.
The five decisions, and what each one actually does
These five are the whole option set. The table separates three things that get confused: whether Google still fetches the URL, whether it can appear in results, and what happens to any ranking signal the URL had. The crawling column comes from Google’s large-site crawl budget guide, which is the only place these mechanics are spelled out plainly.
| Decision | Does Google still crawl it | Can it appear in search | What happens to its signals | Use it when |
|---|---|---|---|---|
| Leave it alone | Yes | Yes | Unchanged | The page already answers the query it ranks for. Most of a healthy inventory sits here. |
| Improve it in place | Yes | Yes | Kept, and the URL keeps its history | The topic is right and the execution is thin. This is the decision Google’s guidance points at first. |
| Consolidate into another page | The old URL is requested less once a redirect is in place | The destination appears, not the source | Passed to the destination | Two or more pages compete for one intent. Google’s crawl guide says to "eliminate duplicate content to focus crawling on unique content rather than unique URLs". |
| Keep the page, remove it from search with noindex | Yes, and that is the catch | No | Lost, while the crawl cost stays | Humans need the page and searchers do not. Google is explicit: "Don’t use noindex, as Google will still request, but then drop the page when it sees a noindex meta tag or header in the HTTP response, wasting crawling time." |
| Delete it and return 404 or 410 | Eventually no. "A 404 status code is a strong signal not to crawl that URL again." | No | Lost permanently | Last resort, and only when the content cannot be salvaged, in Google’s own words. |
Two rows deserve a second look. The noindex row is the one audits get wrong most often, because noindex feels like cleanup and is not: the URL keeps costing crawl requests while contributing nothing, which is the worst of both outcomes if the real goal was to reduce waste. And there is a sixth option people reach for that is not on this list, because it does not do the job: blocking a URL in robots.txt. Google’s robots rules documentation states that "if a page is disallowed from crawling through the robots.txt file, then any information about indexing or serving rules will not be found and will therefore be ignored", so a disallowed URL that is already indexed can simply stay indexed. One more trap worth naming: a deleted page that returns 200 with an empty template becomes a soft 404, and Google says "soft 404 pages will continue to be crawled, and waste your budget".
The original number: where the hours actually pay
The hard part of a content audit is not the decision tree, it is that improving a page costs real hours and a site has thousands of pages. So price it. Take a site with 2,000 content URLs and 40,000 organic clicks a month, with clicks distributed the way they usually are: heavily concentrated at the top. Assume a serious rewrite of one page takes six hours, including research, drafting and internal linking. The last column is the one that decides your audit.
| Band of the inventory | URLs | Share of clicks | Monthly clicks | Clicks per URL | Hours to rewrite the whole band | Hours per monthly click defended |
|---|---|---|---|---|---|---|
| Top 100 | 100 | 60% | 24,000 | 240.0 | 600 | 0.03 |
| Next 300 | 300 | 25% | 10,000 | 33.3 | 1,800 | 0.18 |
| Next 600 | 600 | 13% | 5,200 | 8.7 | 3,600 | 0.69 |
| Bottom 1,000 | 1,000 | 2% | 800 | 0.8 | 6,000 | 7.50 |
| Whole inventory | 2,000 | 100% | 40,000 | 20.0 | 12,000 | 0.30 |
Read the right-hand column as a price list. Rewriting the top 100 pages costs about two hundredths of an hour per monthly click those pages already carry. Rewriting the bottom 1,000 costs 7.50 hours per monthly click, which is 300 times worse. Put the same way: rewriting the tail is 6,000 hours of work defending 800 clicks a month, and those 6,000 hours would buy 60 hours of attention for every page in the top 100, which carry 24,000.
The distribution does the rest of the arguing. Half the inventory is carrying 2 percent of the clicks. The top 20 percent is carrying 85 percent. An audit that spends its budget evenly across 2,000 URLs has spent most of it on the half that cannot repay it.
What this model gets wrong
Three things, and they matter enough to state before anyone quotes the 300.
- It prices defending existing clicks, not winning new ones. A tail page with 0.8 clicks a month might be one rewrite away from 200, and the model cannot see that. Potential, judged from impressions and average position, belongs alongside clicks in any real inventory.
- It treats pages as independent. They are not. Google’s core updates guidance frames the assessment at the level of "your site as a whole", and a large block of pages nobody wants can affect how the site is assessed, which is a site-level cost the per-page arithmetic does not capture.
- Six hours per rewrite is a single number standing in for a wide range. A product page refresh and a researched guide are not the same job. The ratios survive a different constant; the absolute hours do not.
The audit sequence I run
- Build the inventory from a crawl of the site, not from the CMS. The CMS lists what you published; the crawl lists what Google can reach.
- Join it to Search Console clicks, impressions and average position per URL over at least the last 12 months, so seasonality does not condemn a page.
- Join it to the Page Indexing report. A page with no clicks because it is not indexed is a different problem from a page that is indexed and ignored, and they have different fixes.
- Find the competing sets first. Two or more URLs ranking for the same query is the one finding where consolidation reliably adds rather than redistributes.
- Band the remainder by clicks and by impressions, and compute the hours-per-click figure above with your own numbers. That is your spending plan.
- Apply Google’s own self-assessment questions to the head, page by page. Original analysis, clear sourcing, and a reason to exist that is not ranking.
- For the tail, default to leaving it alone. Reach for noindex only when humans need the page, knowing it keeps costing crawl requests, and for deletion only when the content genuinely cannot be salvaged.
- Record the decision and the date for every URL, then measure at 30, 90 and 180 days. Google’s own timeline runs to several months, so a verdict at four weeks is noise.
The deliverable that comes out of this is not a delete list. It is a short, costed list of pages to improve, a shorter list of sets to consolidate, and an explicit decision to leave most of the inventory alone. That is a less impressive document than a spreadsheet with a thousand red rows, and it is the one that matches what Google has actually published.
Related reading
- Google penalty recovery: how I diagnose and reverse a traffic collapse
- Thin content: there is no word count and no thin content penalty
- The SEO audit checklist I actually use, ordered by what moves traffic
- X-Robots-Tag: the header that reaches the files a meta tag cannot
- Index bloat: what it costs in recrawl days
Leave a Reply