Technical SEO Audit: A Step-by-Step Process That Puts Ranking Blockers First

A technical SEO audit checks whether search engines can discover, crawl, render and index a website's pages, and whether those pages load well enough for real users. It covers robots.txt, status codes, redirects, canonical tags, sitemaps, JavaScript rendering, site architecture, Core Web Vitals, structured data and hreflang. Content quality and backlinks belong to the wider SEO site audit.

This guide follows the order in which Google processes a page, so problems that block everything downstream are found first.

What is a technical SEO audit?

A technical SEO audit is a review of the infrastructure that lets search engines access and understand a website: crawl controls, server responses, indexing directives, canonicalisation, rendering, site structure, performance and markup. Its output is a list of fixes ranked by how much search visibility each issue is costing.

In scopeOut of scope (covered by a full site audit)
robots.txt, meta robots, X-Robots-TagKeyword targeting and content gaps
Status codes, redirects, soft 404sContent quality and E-E-A-T
Canonical tags and duplicate URLsBacklink profile
XML sitemapsCompetitive gap analysis
JavaScript rendering, <head> validityBrand and AI citation share
Click depth, orphan pages, internal links
Core Web Vitals, HTTPS, mobile parity
Structured data validity
hreflang
Server logs and crawl budget (large sites)

When to run one

The often-quoted "every six months" is reasonable for active sites, but the triggers above matter more than the calendar.

Tools you need

A technical audit doesn't need an expensive stack:

ToolJobCost (checked Sep–Oct 2026)
Google Search ConsoleIndexing reasons, URL Inspection, crawl stats, Core Web Vitals field dataFree
PageSpeed Insights / LighthouseField and lab performance dataFree
Rich Results Test / Schema Markup ValidatorStructured data validityFree
Screaming Frog SEO SpiderDesktop crawler with JavaScript renderingFree up to 500 URLs; £199 per licence per year (1–4 licences)
Ahrefs Site AuditCloud crawlerFree for verified sites on the free plan; Lite from $129/month
Semrush Site AuditCloud crawler100 pages/month free; Pro $139.95/month or $117.33/month billed annually
Server log filesWhat Googlebot actually requestedFree (from your host or CDN)

Prices change; confirm on the vendor's pricing page before budgeting.

How to conduct a technical SEO audit

How Google processes a page Audit steps 1–13, mapped to the stage each one tests Discover Sitemaps, site architecture, AI crawler rules Steps 6, 8, 13 Crawl robots.txt, redirects, logs & crawl budget Steps 1, 2, 5, 12 Render JavaScript & raw-vs-rendered DOM Step 7 Index Directives, canonicals, schema, hreflang Steps 3, 4, 10, 11 Serve Core Web Vitals & page experience Step 9

The search pipeline — Discover, Crawl, Render, Index, Serve — with the 13 audit steps mapped to the stage each one tests.

1. Crawl as Googlebot smartphone

Google indexes the mobile version of your site, so configure your crawler accordingly:

Search Console URL Inspection confirming a page was crawled as Googlebot Smartphone
Confirm Google is actually crawling your mobile version — URL Inspection shows exactly what it crawled as.

2. robots.txt and server responses

Check that /robots.txt:

Then open Search Console → Settings → Crawl stats and look at responses by type. A rising share of 5xx responses or slow average response times is a server problem, not an SEO tweak.

Google Search Console crawl stats report grouped by response code, with 5xx and robots.txt rows highlighted
Server errors here can stall crawling of the whole site.

3. Indexability directives

List every URL with noindex (meta robots or X-Robots-Tag header) and confirm each is intentional. Then check for these conflicts:

4. Canonicalisation

For each indexable template, check that canonical tags:

Then sample important URLs in Search Console's URL Inspection and compare User-declared canonical with Google-selected canonical. A mismatch means Google has overruled you, usually because other signals point elsewhere. The canonical tags guide covers the five most common mistakes.

5. Status codes and redirects

From the crawl, list:

More on redirect choice in 301 redirects.

6. XML sitemaps

A sitemap should contain only URLs you want indexed: 200 status, indexable, self-canonical. Check:

A clean sitemap turns Search Console's indexing report into a diagnostic: if a URL is in the sitemap and not indexed, you know it's a problem. See XML sitemaps.

7. Rendering and JavaScript

Compare the raw HTML (view-source, or the crawler's "original HTML") with the rendered HTML for each major template. Check that these exist in the raw HTML, or at minimum in the rendered HTML Google sees in URL Inspection:

Also check for invalid elements in <head>. If an <img>, <div> or <a> appears in the head, browsers and parsers can end the head early and treat everything after it as body content, where title, canonical and robots tags no longer work. About 10% of pages have this problem, per the 2025 Web Almanac, and it's often caused by a tag manager or third-party script injected high in the page.

Side-by-side comparison of raw HTML and rendered HTML showing a canonical tag added only by JavaScript
If a tag only exists after JavaScript runs, Google may not see it on the first pass.

From the crawl:

Internal linking for SEO covers anchor text and fixing orphan pages.

9. Core Web Vitals and page experience

Use field data (Chrome UX Report, via Search Console's Core Web Vitals report or PageSpeed Insights) grouped by template. Google's "good" thresholds at the 75th percentile:

Bar chart showing Core Web Vitals pass rates by metric and device, June 2025
Mobile LCP is the most common failure. Source: 2025 Web Almanac.

Also confirm HTTPS on every URL (91.7% of desktop pages use it), no mixed content, and that the mobile page contains the same content, links and structured data as desktop. Google's Mobile-Friendly Test was retired in December 2023; use Lighthouse for mobile checks. Detail in the Core Web Vitals guide.

10. Structured data

Validate markup with the Rich Results Test and the Schema Markup Validator. Check:

11. International and hreflang

For multilingual or multi-regional sites:

The hreflang tags guide lists the five errors that break it.

12. Log files and crawl budget (large sites only)

Server logs show what Googlebot actually requested, how often and with what response. They reveal crawl spent on parameters, redirects and 404s, and important sections Googlebot rarely visits.

Google's own guidance says crawl budget is a concern for sites with over a million unique pages changing about weekly, or over 10,000 pages changing daily. Below that, crawl-budget optimisation rarely moves results. Fix duplicates and soft 404s for hygiene and move on.

13. AI crawler access

Review robots.txt rules for AI crawlers such as GPTBot, ClaudeBot and PerplexityBot, and the Google-Extended token (which governs use of content for Google's AI models, not Google Search crawling). Rules for these grew quickly: GPTBot was named in 4.5% of desktop robots.txt files in 2025, up from 2.9% in 2024, and ClaudeBot in 3.6%, up from 1.9%.

Allowing or blocking them is a business decision. The audit finding is when the current rules weren't chosen deliberately, for example a security plugin blocking every AI crawler by default on a site that wants to be cited in AI answers.

On llms.txt: about 2% of sites have a valid one, and Google has said it doesn't use the file. Treat it as optional; see llms.txt explained.

How to rank technical issues by severity

Crawlers label issues as errors, warnings and notices. Those labels describe technical correctness, not business impact. Use severity instead:

SeverityDefinitionExamples
CriticalStops pages being crawled, rendered or indexed at scaleSitewide noindex or robots.txt block; robots.txt returning 5xx; main content only visible after JavaScript and not rendered; canonical on every page pointing to the homepage
HighCosts visibility on important pages or templatesGoogle-selected canonical differs on money pages; key templates failing Core Web Vitals; redirect chains on high-traffic URLs; orphaned category pages; hreflang errors on main markets
MediumWastes crawl or weakens signalsInternal links to 404s; 302s for permanent moves; noindexed URLs in sitemap; invalid structured data on secondary templates
LowHygiene; fix when convenientMissing alt text on decorative images; long title tags; minor HTML validation warnings

Then order within each band by the traffic or revenue the affected pages carry. A medium issue on your top category can outrank a high issue on an archive nobody visits.

Writing findings developers will act on

A technical SEO audit report is only as useful as its weakest finding. Write each one in the same structure:

FieldExample
FindingPaginated category pages canonicalise to page 1
SeverityHigh
AffectedCategory template — 312 URLs (/category/*?page=2+)
EvidenceURL Inspection on /shoes?page=3: Google-selected canonical = /shoes; products on pages 2+ "Discovered – currently not indexed"
Why it mattersProducts listed only on deeper pages lose their main internal link path
FixMake each paginated page self-canonical; keep <a href> links between pages
Effort~1 day (template change)
How to verifyRe-inspect 5 sample URLs after deploy; watch Page indexing for affected products over 4–6 weeks

Findings written this way can go straight into a ticketing system.

If you'd rather have this done for you, the Technical-Only Deep Dive delivers the full technical review with code-level fixes, and the Site SEO Audit adds content, links and AI visibility.

Frequently asked questions

What is included in a technical SEO audit? Crawl controls (robots.txt, meta robots), server responses and redirects, canonical tags, XML sitemaps, JavaScript rendering, site architecture and internal links, Core Web Vitals, HTTPS, structured data, hreflang, and for large sites, log file and crawl budget analysis.

How long does a technical SEO audit take? A site with a few hundred pages can be audited in two to three days. Large or JavaScript-heavy sites take longer, mainly for rendering analysis, log processing and writing findings developers can implement.

Can I do a technical SEO audit myself? Yes. Search Console, PageSpeed Insights, the Rich Results Test and Screaming Frog's free 500-URL crawl cover most checks on a small site. The hard parts are interpreting rendering and canonical conflicts, and deciding what matters most.

How often should you run a technical SEO audit? Annually as maintenance, and always before and after migrations, redesigns or platform changes. Monitor Search Console's Page indexing and Core Web Vitals reports monthly between audits.

Is crawl budget something every site should audit? No. Google says it matters mainly for sites with over a million pages, or over 10,000 pages changing daily. Smaller sites are generally crawled efficiently.

What's the difference between a technical SEO audit and an SEO audit? A technical audit checks whether search engines can access and process the site. A full SEO audit adds content, keyword targeting, backlinks, competitors and AI visibility on top.

Get your technical SEO audit done for you

Running all thirteen steps correctly takes real crawler expertise — reading a raw-vs-rendered HTML diff, telling a harmless robots.txt 404 from a dangerous 5xx, and ranking findings by revenue impact rather than crawler-label severity. TapasSEO's Technical-Only Deep Dive runs the full pipeline above against your site and hands back a prioritised, developer-ready findings list — not a list of errors with no order.

Technical-Only Deep Dive — $697 →

Last updated 2026-10-05 · Written by .