Phase 0 · DecideModule 2 of 15

Find out where you actually stand

Most audits produce a four-hundred-row spreadsheet nobody acts on. The useful version fits on one page. Here is the one I run, including on my own site, where it did not go well.

What this actually costs you

Hours to learn
4
Hours / month
1
Tools
$0 to $139/mo
Difficulty
Low to run, high to interpret

What this looks like when it is done badly. Somebody runs a site crawler, exports 412 issues, sorts by the tool’s own severity column, and spends a weekend fixing missing meta descriptions on pages that are not indexed. The report gets filed. Nothing about the site changes.

Hours and tool costs are RedSEO’s own estimates from client work and from teaching this material, not industry survey data. They assume a business owner doing the work themselves on a site under about 50 pages.

Almost every audit I have been sent by a prospective client has the same shape. Four hundred rows, exported from a crawler, sorted by that crawler’s idea of severity. Nobody has read past row forty. Nothing in it says which pages make money, so nothing in it can say which issues matter.

A baseline is not that. A baseline is the short list of numbers you will judge every later month against, and it fits on one page. Four numbers, an afternoon, and no budget.

Your baseline is four numbers

In the order they matter:

  1. What share of your pages are actually indexed. Not how many pages you have. What fraction of them Google has agreed to store.
  2. What you already get impressions for. Especially on page two, where you are one push away from traffic.
  3. Clicks over a fixed 28-day window, with the dates written down.
  4. One conversion you can trace to organic. One is enough to start.

Everything else in a standard audit is a means to changing one of those four. If a finding cannot be connected to one of them, it goes at the bottom of the list, whatever colour the tool printed it in.

Start with the source that is free and authoritative

Google Search Console “provides information on how Google crawls, indexes, and serves websites,” in Google’s own words. That distinction matters more than it sounds. Every third-party tool is inferring what Google did. Search Console is Google telling you. It costs nothing and it is the only source in this module that is not a guess.

Two reports carry the whole baseline. The Performance report shows “how much traffic you’re getting from Google Search, including breakdowns by queries, pages, and countries.” The Page Indexing report “gives you an overview of all the pages Google indexed or tried to index in your website.” Google suggests checking “around once every month,” which is also roughly the cadence at which anything you change will have had time to register.

Verify with the domain property rather than a URL prefix if you can add a DNS record. A URL-prefix property silently excludes the other protocol and every subdomain, which is how people end up auditing half their site without knowing it.

The number almost nobody checks

Here is the one that changes what people do next. Google’s own documentation is blunt about it: “Google doesn’t guarantee that it will crawl, index, or serve your page, even if your page follows the Google Search Essentials.” Crawling, indexing and serving are three separate stages, and Google says plainly that “not all pages make it through each stage.”

So the useful question is not how many pages you have. It is what percentage of them survived to the index. Divide indexed pages by total pages and you have an indexation rate, and once you have it as a percentage rather than a count you can compare parts of your site against each other.

I will use my own site, because it makes the point better than a client’s would. In August 2026 I split redseo.io’s sitemap into segments specifically so I could measure this, and the answer was not good.

Figure 01Our own auditredseo.io indexation rate, by section
redseo.io indexation rate, by sectionBar chart of indexation rate by sitemap segment on redseo.io, measured August 2026: blog 37 percent, glossary 68 percent, service areas 81 percent, tools 86 percent, and compare, services and templates at 100 percent.Blog37%Glossary68%Service areas81%Tools86%Compare, services, templates100%

The worst-performing part of my own site is the blog, at 37 percent. Fifty-nine of ninety-three posts are outside the index. A single site-wide average would have put this in the sixties and hidden which half of the site was the problem.

RedSEO’s own sitemap segments measured against Search Console exclusions, 21 August 2026

The blog is the worst-performing part of my own website by a wide margin. Thirty-seven percent. Fifty-nine of ninety-three posts sitting outside the index, which is fifty-nine posts of writing that cannot rank for anything because they are not eligible to. The tools, which get a fraction of the attention, are at eighty-six percent.

I would not have found that with a crawler, because a crawler tells you about the pages, not about what Google did with them. And I would not have found it with a single site-wide number either, because averaging 37 percent against 100 percent produces something in the sixties that looks survivable and hides which half of the site is the problem. Segment the measurement and the answer tells you where to work.

Have us measure your indexation rate properly

Reading the Not indexed list without panicking

The Page Indexing report groups un-indexed pages by reason, and the reasons are not equally serious. Roughly, in the order I triage them:

  • Excluded by ‘noindex’ tag. You did this, on purpose or by accident. Check whether you meant it. This is the one that occasionally reveals that a staging setting shipped to production.
  • Duplicate without user-selected canonical. Google found several versions of the same thing and picked one for you. That is an architecture problem, and it is module 4.
  • Crawled, currently not indexed. Google looked and decided not to store it. Uncomfortable, because there is no technical fix. Google lists the causes as “the quality of the content on page is low” among others. This is a content problem wearing a technical costume.
  • Discovered, currently not indexed. Google knows the URL exists and has not got to it. On a small site this usually resolves. On a large one it is the beginning of a crawl budget conversation, which module 4 will talk you out of.

What a crawler adds, and what it does not

Once you have the Search Console picture, a crawler earns its place by telling you things Google will not: which pages have no internal links pointing at them, where your title tags are duplicated, how deep your click depth gets. That is genuinely useful and it is what the audit tool below does.

What it cannot do is prioritise. A crawler does not know that your services page is worth two hundred times your privacy policy, so its severity column is sorted by a scale that has nothing to do with your business. The triage rule that actually works: does fixing this change one of the four numbers, on a page that matters? Most findings do not, and saying so out loud is most of the skill.

What I would do in your position

Give it one afternoon. Verify Search Console, write the four numbers on one page, date it, and stop. Do not fix anything yet. The temptation to start fixing while you are still measuring is strong and it is how you end up unable to tell whether anything worked.

Then look at exactly two things. Your indexation rate by section, because that tells you whether you have a publishing problem or an eligibility problem. And your page-two queries, because that is where the cheapest traffic in your entire SEO effort is sitting. Both are free. Neither requires you to buy anything or to have finished this curriculum.

If the indexation rate is low, the next module you want is not the next module in order. Go to module 7 and check the indexing basics, then come back. If the rate is fine and the page-two list is long, you are in better shape than you think and module 5 will pay for itself quickly.

Run it on your own site

The site audit takes a URL and does the crawl for you. The rest are one-input tools that answer a single question each, which is all you need at this stage.

Crawl your site

Enter your domain and the audit will crawl it, score the on-page basics, and list what it finds. Read the output with the triage rule above rather than top to bottom.

Opens the full audit on the tool page. Larger sites take longer to crawl.

Open the SEO Site Audit on its own page

The three other numbers

One question each. Together with Search Console they give you the whole baseline.

  • Website Page CounterHow many pages you actually have, which is the denominator of your indexation rate.
  • Domain KeywordsWhat you already rank for, including the page-two terms worth going after first.
  • AI Visibility CheckerWhether assistants mention you at all. A separate baseline, and module 12’s problem.

Write down your baseline in one afternoon

The output is one page with four numbers on it and a date. If it takes longer than an afternoon, you have started fixing things, which is the next module.

  1. Verify your site in Google Search Console

    Free, and it is the only source that tells you what Google actually did with your site rather than what a third party guessed. Use the domain property if you can add a DNS record, because it covers every subdomain and protocol at once.

  2. Write down clicks and impressions for the last 28 days

    Performance report, not Analytics. This is the number every later month gets compared against, so record the date range next to it or the comparison is meaningless.

  3. Count how many pages you have, and how many are indexed

    Pages you have comes from your sitemap or a crawl. Pages indexed comes from the Page Indexing report. Divide the second by the first. That percentage is the single most useful number in this module.

  4. Read the top reasons in the Not indexed list

    Group them. "Crawled, currently not indexed" is a quality signal. "Excluded by noindex tag" is something you did. "Duplicate without user-selected canonical" is an architecture problem, which is module 4.

  5. List every query you already get impressions for on page two or three

    Filter the Performance report to positions 11 to 30. These are the cheapest wins you will ever have, because Google already thinks you are nearly good enough.

  6. Check that your money pages are actually in the index

    Not your blog. The pages that make you money. Use the URL Inspection tool on each one individually. It takes two minutes and it is the check that most often finds something.

  7. Connect one conversion to organic traffic

    A form fill, a call, a booking, a sale. One is enough to start. If you cannot connect any conversion to a source, that is your baseline finding, and it comes before any SEO work at all.

  8. Date the page and put a reminder in for 90 days

    A baseline you never compare against is just a document. The comparison is the entire point of writing it down.

Can you do this yourself?

You can do this yourself if

  • You can read a chart and resist the urge to act on every red line in it. Most audit output is noise, and the skill is triage rather than detection.
  • Your site is small enough to hold in your head. Under about fifty pages, you can check the important ones by hand and be genuinely thorough.
  • You are willing to write down a number you do not like and leave it written down.

You cannot do this yourself if

  • You cannot interpret what the output means, which is the normal case. A crawler will hand you four hundred issues sorted by its own severity scale, and that scale does not know which of your pages make money.
  • Your Not indexed list is dominated by "Duplicate without user-selected canonical" or "Alternate page with proper canonical tag". Those are architecture problems, and fixing them wrongly makes things worse rather than nothing.
  • You have no analytics history at all and need a defensible before-and-after for someone else, a lender, an investor, or a buyer, on a deadline you do not control.

Questions people ask about this

What should an SEO audit checklist actually cover?
For a baseline, four things: what share of your pages are indexed, what queries you already get impressions for, clicks over a fixed 28-day window, and one conversion traced to organic. Everything else in a standard audit exists to move one of those four. If a finding cannot be connected to one of them, it is not urgent.
How do I check if my pages are indexed by Google?
Use the Page Indexing report in Search Console for the site-wide picture, and the URL Inspection tool for individual pages. Google is explicit that it "doesn't guarantee that it will crawl, index, or serve your page", so a page existing on your site is not the same as a page being eligible to rank.
What is a good indexation rate?
Higher than yours probably is, and the more useful version of the question is by section rather than site-wide. When we measured redseo.io in August 2026 the blog was at 37% while the tools were at 86%, and the site-wide average hid that completely. Compare your sections against each other before you compare yourself against anyone else.
Do I need a paid tool to audit my own site?
No. Search Console is free and it is the only source that reports what Google actually did rather than inferring it. A crawler adds things Google will not tell you, like orphaned pages and duplicate titles, and the free site audit here covers that. Paid tools become worth it at scale, not at the baseline stage.
How long does a baseline audit take?
An afternoon, if you resist fixing things while you measure. Our estimate is four hours to learn the reports and about an hour a month to keep the baseline current. The honest cost table prices every module the same way.
What is a baseline ranking check?
A dated record of where you rank for your target terms before you change anything, so later movement is attributable. It is the ranking half of the baseline; the other half is indexation and conversions. We wrote about it separately in what is a baseline ranking check.