How to Conduct a Technical SEO Site Audit: A Practical Checklist

To conduct a technical SEO site audit, crawl the site and compare the results with Google Search Console. Then assess crawlability, indexation, architecture, duplicates, speed, mobile usability, HTTPS, and schema before prioritizing fixes by visibility risk and revenue impact.
How to Conduct a Technical SEO Site Audit: A Practical Checklist
Picture of Peter Strauss

Peter Strauss

Peter Strauss is an eCommerce SEO specialist with over eight years of experience driving organic growth for digital brands. Specializing in Shopify, WooCommerce, and WordPress environments, he blends technical architecture optimization with revenue-focused content strategy to help online retailers capture market share and scale sustainably.
Picture of Peter Strauss

Peter Strauss

Peter Strauss is an eCommerce SEO specialist with over eight years of experience driving organic growth for digital brands. Specializing in Shopify, WooCommerce, and WordPress environments, he blends technical architecture optimization with revenue-focused content strategy to help online retailers capture market share and scale sustainably.

Organic growth usually stalls long before a team runs out of content ideas. Knowing how to conduct a technical SEO site audit gives you a way to find the infrastructure problems quietly keeping collection pages, product pages, and buying guides out of search results.

The old approach is to export hundreds of crawler warnings, attach a health score, and call it an audit. High-performing teams do something more useful: they isolate root causes, measure which templates and revenue paths are affected, and hand developers a sequence of fixes that can be verified after release.

For an eCommerce operator, this is commercial hygiene. A misconfigured filter system can create thousands of low-value URLs, while one accidental noindex tag on a collection template can remove a major organic acquisition path overnight.

How to Conduct a Technical SEO Site Audit: Start With the Right Scope

How to Conduct a Technical SEO Site Audit: Start With the Right Scope

A technical SEO site audit is a systematic review of the factors that determine whether search engines can crawl, render, index, and understand your website. It checks crawlability, indexability, architecture, speed, mobile usability, security, and structured data before prioritizing the fixes with the greatest search impact.

The point is not to produce a longer issue list. The point is to identify barriers preventing valuable pages from being found, interpreted correctly, and surfaced to qualified buyers.

What a Technical SEO Audit Covers and What It Does Not

Technical SEO is the operating layer beneath your search visibility. It examines access controls, status codes, internal paths, duplicate URLs, canonical signals, rendering, performance, mobile behavior, security, metadata, and structured data.

A broader SEO audit also evaluates content quality, keyword targeting, backlinks, brand authority, and conversion opportunities. Those matter, but they should not distract from a technical site audit when Google cannot consistently reach or index the pages meant to generate organic revenue.

The core pillars are:

  • Crawl access: Can bots reach essential HTML, resources, and URLs?
  • Index selection: Are the right canonical pages eligible for indexing?
  • Site architecture: Can users and crawlers reach priority pages through logical internal links?
  • Template health: Are speed, mobile, schema, and metadata working reliably at scale?
  • Signal consistency: Do redirects, canonicals, sitemaps, and internal links agree?

For a useful technical SEO audit, treat crawlability and indexability as separate checks. A URL may be crawlable but excluded from the index, or blocked from crawling while still appearing in reports because search engines discovered it elsewhere.

Set Your Audit Goal, Scope, and Baseline Before You Crawl

A crawl without a business baseline creates noise. Before opening a crawler, document the pages and templates that matter most: top organic landing pages, high-impression collections, best-selling products, seasonal categories, and pages with a recent traffic decline.

Record organic clicks, impressions, sessions, conversions, revenue, indexed URL estimates, Core Web Vitals status, and recent releases. Include theme changes, migrations, redirects, catalog imports, new apps or plugins, and merchandising changes, because timing often exposes the cause.

A store with 30,000 URLs does not necessarily have a 30,000-URL problem. It may have a single collection-filter template generating 25,000 parameter variations. Scope the SEO audit around patterns, not raw URL volume.

Step 1: Gather Your Technical SEO Audit Tools and Data

One data source is never enough. Google Search Console shows Google’s view of indexing and search performance, while a crawler exposes sitewide patterns that may not yet appear in Search Console.

Analytics adds the business context. A 404 on an abandoned blog post is not equivalent to a 404 on a collection that generated 10 percent of organic revenue last month.

Use tools to surface clues, then validate them. Automated issue counts cannot tell you whether a warning is intentional, whether a developer deployment caused it, or whether it affects pages that buyers actually use.

The Minimum Tool Stack for Small and Growing eCommerce Stores

The minimum reliable stack is Google Search Console, GA4 or equivalent analytics, a crawler, page-performance testing, and manual browser checks on real devices. Smaller stores can begin with free checks and targeted crawls; larger catalogs need more structured crawling and exports.

Google Search Console helps you review Page Indexing, sitemaps, URL Inspection, performance queries, Core Web Vitals, and security problems. A crawler such as Screaming Frog, Sitebulb, or a Site Audit crawl helps reveal status-code patterns, canonicals, directives, internal-link gaps, duplicate metadata, and URL depth.

Use PageSpeed Insights or Lighthouse for representative product, collection, article, and cart-adjacent pages. Then test the same routes manually on a mobile device. A desktop score cannot reveal a filter drawer that covers the add-to-cart button.

Configure the Crawl Before You Press Start

Default settings miss important evidence. Configure the crawl to collect XML sitemap URLs, status codes, title tags, meta tags, canonicals, meta-robots directives, internal links, crawl depth, images, pagination, and structured data.

Enable JavaScript rendering when the store relies on client-side navigation, injected product content, or app-based filtering. If the site is heavily script-driven, run a second comparison between rendered and non-rendered output. That contrast often reveals links or product details that only exist after a browser executes JavaScript.

For Shopify or WooCommerce stores, a technical audit should account for product, collection, filter, and app-generated URLs, not just generic site-health warnings. SEO.DIGITAL’s eCommerce-focused audits are built around those patterns.

Step 2: Check Crawlability and Indexability First

Step 2: Check Crawlability and Indexability First

Do not polish title tags on pages search engines cannot access or do not want to index. Crawlability and indexability come first because every other improvement depends on priority URLs being discoverable and eligible for search.

Run this first-pass checklist before moving on:

  1. Review the robots.txt file and important meta-robots directives.
  2. Compare the XML sitemap with URLs that should rank.
  3. Inspect Page Indexing statuses in Google Search Console.
  4. Test representative URLs with URL Inspection.
  5. Check status codes, canonicals, and rendered content together.

Audit the robots.txt File and Meta-Robots Directives

A single Disallow rule can block a valuable site section. Check the robots.txt file for staging directives left behind after launch, blocked product or collection paths, and blocked CSS or JavaScript resources needed to render meaningful content.

Robots.txt controls crawl access. A noindex meta tag controls index eligibility after a crawler can access the page. They are not interchangeable, and using both without a clear reason creates confusing signals.

Review important templates for noindex, nofollow, and X-Robots-Tag directives. Common failures include a noindex rule inherited by collection pages, products marked noindex after they return to stock, or a development setting deployed to production.

Review the XML Sitemap Against the URLs That Should Rank

Your XML sitemap is a declaration of preferred, indexable URLs. It should not contain redirects, 404s, server errors, noindex pages, duplicate parameter URLs, or canonicals pointing elsewhere.

Export sitemap URLs and compare them with crawl data. On an eCommerce store, investigate outdated product URLs, filtered collection URLs, tag archives, discontinued products, and variant paths that have slipped into the sitemap.

A sitemap is not a dumping ground for every URL the platform can generate. Keep it focused on pages with search value and a consistent canonical signal.

Use Google Search Console to Verify Actual Indexation

Crawlers show what the site exposes. Google Search Console shows how Google has classified at least part of that exposure. Compare both views before escalating an issue.

Review Page Indexing statuses in context:

  • Crawled, currently not indexed: Investigate content value, duplication, rendering, canonical signals, and internal links.
  • Discovered, currently not indexed: Check crawl demand, internal linking, sitemap quality, server performance, and large-scale URL waste.
  • Blocked by robots.txt: Confirm whether the block is intentional and whether the URL should appear in search.
  • Alternate page with proper canonical: Often normal, if the chosen canonical is correct and indexable.
  • Duplicate without user-selected canonical: Investigate competing URL versions and strengthen the preferred signal.

Use URL Inspection on a small set of high-value pages from each affected template. You are looking for a pattern, not trying to manually inspect every product.

Step 3: Audit Site Architecture, Internal Linking, and URL Paths

A strong product page buried six clicks deep is hard for users to find and easy for crawlers to treat as low priority. Architecture decides how authority, discovery, and shopper attention move through the site.

Prioritize collection, category, product, and evergreen buying-guide pages with meaningful impressions, conversions, or commercial intent. Navigation should support both shopping behavior and search discovery.

Measure Crawl Depth and Find Orphan Pages

Important pages should generally be reachable within about three clicks from the homepage. That is guidance, not a universal law, but product or collection pages sitting four or more clicks deep deserve scrutiny.

Review crawl-depth reports for pages linked only through the footer, internal search, or long pagination paths. Then compare crawler URLs against sitemaps, analytics landing pages, and Search Console data to identify orphan pages that receive no internal links.

A practical eCommerce path is simple: Home → Collection → Subcollection → Product. If a product is only reachable through a filtered search result, its discoverability and internal-link equity may be weak.

Broken internal links force both users and crawlers into dead ends. Prioritize errors on navigation, breadcrumbs, category modules, related-product widgets, and high-traffic editorial pages.

For each 404 or 5xx response, decide whether to restore the page, update the internal link, create a direct 301 redirect to the closest equivalent page, or retain a useful 404 because no equivalent exists. Redirecting every missing product to the homepage is rarely useful to shoppers or search engines.

Also find internal links pointing to redirects. Replace them with the final destination. Redirect chains and loops add latency, obscure intent, and make troubleshooting much harder when a release goes wrong.

Review URL Structure, Pagination, Filters, and Faceted Navigation

URL patterns should communicate a stable site architecture. Clean descriptive paths are easier to manage than uncontrolled parameters, but the biggest issue is usually not cosmetic URLs. It is uncontrolled URL generation.

Filters for color, size, brand, price, sort order, and availability can create vast combinations. Product variants, tracking parameters, pagination, and search-result URLs can add even more duplicate content and crawl waste.

Do not apply one universal rule to faceted navigation. A filter page with distinct demand and useful inventory may deserve an indexable, curated landing page. Low-value combinations may need controlled crawl paths, canonical handling, noindex logic, or removal from internal discovery. Make the decision based on unique search value, not fear of parameters.

Step 4: Diagnose Duplicate Content and Canonical Tag Issues

Step 4: Diagnose Duplicate Content and Canonical Tag Issues

Duplicate content is usually a signal-clarity problem, not a penalty story. When several near-identical URLs compete, search engines must decide which one represents the page, which can dilute internal signals and produce unpredictable index selection.

eCommerce sites create these conflicts naturally. The answer is consistent URL governance, not indiscriminate canonical tags.

Identify Common Duplicate URL Patterns

Look for multiple versions caused by HTTP and HTTPS, www and non-www hosts, trailing slashes, index files, tracking parameters, parameter order, sorting, print views, filters, variants, and pagination.

A product may be visible through its direct URL, a collection path, a variant parameter, and multiple campaign-tagged URLs. A collection may appear in dozens of filtered states with nearly identical copy and products.

Check whether each pattern is internally linked, included in XML sitemaps, receiving impressions, or collecting backlinks. These clues show which duplicate paths are actively competing rather than merely existing in logs.

Validate Canonicals Instead of Blindly Trusting Them

A canonical tag is a hint. It cannot compensate for incoherent internal linking, weak redirects, blocked target URLs, or a sitemap that promotes competing versions.

For a sample from every key template, verify that the canonical target is live, indexable, relevant, self-consistent, and represented in internal links and the XML sitemap. Flag canonical chains, canonicals to redirected URLs, canonicals to irrelevant pages, and pages marked both canonicalized and noindex.

A classic failure looks like this: filtered collection URLs are canonically pointed to the main collection, but the sitemap still lists the filters and category modules link to them aggressively. The audit must identify that conflict and name the template or app creating it.

Step 5: Audit Page Speed, Core Web Vitals, and Rendering

Slow pages cost attention at the exact moment a buyer is trying to compare products, select a variant, or add an item to cart. Technical SEO performance work should focus on the templates that combine high organic visibility with commercial value.

Do not chase a perfect lab score. Find repeatable bottlenecks across mobile product pages, collection pages, search pages, and media-heavy editorial templates.

Measure Core Web Vitals on Mobile and Desktop

Core Web Vitals give you three useful views of interaction quality. Largest Contentful Paint (LCP) measures when the main visible element loads, Interaction to Next Paint (INP) measures responsiveness after interaction, and Cumulative Layout Shift (CLS) measures unexpected movement.

Use targets of LCP within 2.5 seconds, INP below 200 milliseconds, and CLS below 0.1 as practical benchmarks. Review real-user field data in Search Console alongside lab diagnostics, because a fast test on one device does not erase poor experiences across real shoppers.

Group failing URLs by template. Fifty slow product pages caused by the same hero-image component are one implementation problem, not fifty unrelated tickets. For a closer explanation of thresholds and diagnostics, use this Core Web Vitals guide.

Find the Root Causes of Slow Pages

Performance failures often come from predictable sources: oversized hero images, render-blocking CSS and JavaScript, unused scripts, app or plugin bloat, third-party tags, weak server response, excessive requests, missing caching, and embeds without reserved dimensions.

For LCP, investigate the primary hero image, image delivery format, preload behavior, server response, and critical CSS. For INP, focus on expensive JavaScript tasks, third-party scripts, and interaction handlers around filters, variant selectors, and cart drawers.

For CLS, reserve space for images, banners, reviews, cookie notices, and embedded widgets. A promotional bar that shifts the product price after a shopper arrives is not a minor visual defect. It disrupts purchase intent.

Search engines can render JavaScript, but heavy client-side implementations can delay discovery and introduce inconsistencies. The risk rises when navigation, canonical tags, pricing, product descriptions, or internal links appear only after scripts run.

Compare raw HTML with rendered output for representative templates. Check both desktop and mobile views if navigation changes by device. If a crawler cannot find collection-to-product links without rendering, that is a warning worth investigating.

Critical internal links should be crawlable in HTML whenever possible. Scripted enhancements are fine; hiding the basic information architecture behind fragile JavaScript is not.

Step 6: Test Mobile-Friendliness and HTTPS Security

Mobile issues often hide in the conversion path, not in a generic mobile-friendly design check. A page can technically fit a phone screen while still making filters unusable, product options hard to select, or internal links impossible to tap.

Security belongs in the same audit because inconsistent protocol handling creates browser friction, broken assets, redirect confusion, and trust problems.

Review Mobile UX on Real Conversion Paths

Test the routes shoppers actually take: category navigation, collection filters, onsite search, product media, variant selectors, add-to-cart controls, forms, account pages, and mobile menus. Do this on physical devices when possible.

Look for small tap targets, unreadable text, horizontal scrolling, hidden content, obstructive pop-ups, sticky elements covering purchase controls, and different internal-link behavior from desktop. A mobile menu that excludes key collections can weaken both user discovery and site architecture.

Record the device, browser, template, and reproduction steps. “Mobile usability issue” is not a developer-ready finding. “iPhone Safari, collection filter drawer traps page scroll and hides Apply button” is.

Verify HTTPS, SSL, and Mixed-Content Problems

Confirm that the HTTPS version is valid and that HTTP requests redirect directly to the secure equivalent. Check canonical tags, XML sitemap URLs, hreflang annotations where relevant, and internal assets for protocol consistency.

Mixed-content warnings can occur when a secure page loads an image, script, font, or embed over HTTP. They can break page functionality or trigger browser warnings, especially after app changes and third-party tracking updates.

Review SSL certificate validity and redirect behavior across representative URLs. Protocol consistency is basic search engine optimization hygiene, but basic failures can have wide effects.

Step 7: Validate Metadata, Structured Data, and Page-Level Signals

Template errors multiply quickly. One flawed title-tag rule or stale Product schema field can affect thousands of URLs and make it harder for search engines to interpret pages consistently.

Keep the lens technical. The goal is not to rewrite every page during this audit. It is to identify missing, duplicate, inconsistent, or poorly templated signals that undermine page meaning at scale.

Audit Title Tags, Meta Tags, and Heading Hierarchy

Review title tags and meta tags for missing values, duplication, generic templates, mismatched page intent, and titles that bury the core topic behind boilerplate. Concise titles that front-load the product, category, or primary subject usually communicate relevance more clearly.

Check for missing or duplicate H1s and illogical heading hierarchy. A collection page with several competing H1s, or a product template with no visible primary heading, creates a preventable interpretation problem.

Separate isolated editorial issues from template-wide defects. Five thousand product titles beginning with the same brand boilerplate need a template fix, not five thousand manual edits.

Validate Structured Data Against Visible Content

Structured data helps search engines interpret visible page information, but only when the markup is accurate. Validate JSON-LD against the content a shopper can see, not against what a plugin claims it generated.

Prioritize Product, BreadcrumbList, Article, FAQ, and Review schema where appropriate. On product templates, confirm that name, price, availability, image, rating, and review information match the live page and update when inventory changes.

Structured data is not a promise of a rich result. It is a consistency check that gives eligible pages clearer machine-readable context.

Step 8: Turn Findings Into a Prioritized Technical SEO Roadmap

An audit report fails when it treats all warnings as equal. The next developer sprint cannot absorb 180 disconnected findings, and leadership cannot make a budget decision from a health score alone.

Prioritize root causes by visibility risk, revenue-page impact, template scale, and implementation effort. A single robots block affecting five top collections belongs above 300 missing meta descriptions on low-traffic articles.

Need a roadmap your developers can act on? SEO.DIGITAL turns crawl data, Search Console findings, and eCommerce revenue priorities into a tailored technical SEO action plan.

Use a Four-Tier Prioritization Model

Use four tiers to turn a site audit into an implementation sequence. The tier should reflect evidence and commercial exposure, not how alarming a tool labels the issue.

TierTypical issueDecision criteriaOwner and validation
1Robots blocks, accidental noindex, critical 5xx errors, broken canonicals, widespread 404sVisibility blocker affecting important pagesDeveloper or platform owner; recrawl, inspect URLs, confirm index eligibility
2Filter duplication, redirect chains, weak collection architecture, Core Web Vitals failures on key templatesHigh-value templates or large affected URL groupsDeveloper plus SEO; test template, crawl samples, monitor clicks and impressions
3Schema gaps, metadata patterns, image optimization, orphan pagesScaling or quality improvement with clear upsideSEO, content, or developer; validate template output and coverage
4Low-risk refinements and monitoringLimited exposure or uncertain impactAssigned owner; monitor after higher-risk work is complete

For every issue, capture affected-page count, organic value of those pages, root cause, effort, dependency, owner, and validation method. This prevents teams from spending weeks on a visible but low-impact warning while indexation blockers remain unresolved.

Write Developer-Ready Recommendations Instead of Tool Exports

Developers need reproducible evidence and a clear desired state. Every recommendation should include the issue, affected template or URL examples, evidence, why it matters, root-cause hypothesis, recommendation, priority, owner, effort estimate, and post-release validation.

Consider a collection-filter conflict. A weak ticket says, “Fix duplicate content.” A useful ticket says: “Filtered collection URLs return 200 status, are internally linked by filter controls, appear in the sitemap, and canonicalize to parent collections. Remove filter URLs from the sitemap, confirm intended indexation rules, and update internal links so canonical collection paths receive primary navigation signals.”

That ticket tells engineering what exists, what should change, and how success will be tested. It also gives leadership a reason to prioritize the work.

Re-Crawl, Validate, and Monitor After Fixes

A merged ticket is not a completed SEO fix. Re-crawl affected templates, inspect priority URLs in Search Console, verify status codes and canonicals, and confirm that rendered HTML exposes the expected links and content.

Monitor Page Indexing, Core Web Vitals, organic clicks, impressions, conversions, and revenue after meaningful releases. Rankings can move slowly, but implementation validation should be immediate.

Maintain lightweight monthly monitoring, schedule a fuller quarterly review, and run event-triggered audits after migrations, theme changes, app installations, major catalog updates, or unexplained traffic declines. For keyword and SERP monitoring, you can track page positions and SERP features alongside technical indicators.

A Technical SEO Site Audit Checklist for Shopify and WooCommerce Stores

Shopify and WooCommerce remove many basic operational barriers, but they do not remove technical SEO risk. Theme logic, apps, plugins, catalog structure, product variants, and merchandising workflows can still create large-scale URL, rendering, and performance problems.

Use this compact SEO audit checklist after the core workflow:

  • Product templates: Check canonical logic, variants, Product schema, availability, image weight, title patterns, and out-of-stock handling.
  • Collection templates: Check hierarchy, crawl depth, pagination, filter behavior, internal links, copy, and canonical consistency.
  • URL generation: Review parameter URLs, sorting, search pages, tags, tracking paths, and app-created routes.
  • App and plugin impact: Test script weight, render delays, injected markup, duplicate tags, and conflicts after updates.
  • Mobile conversion paths: Test filters, search, product media, variant selection, add to cart, and checkout-adjacent interactions.

Shopify and WooCommerce Issues to Check First

Begin with platform-specific patterns that have broad template reach. Check duplicate product or variant URLs, collection filtering, pagination, app-injected scripts, internal-link depth, product schema consistency, duplicate title tags, slow theme components, stale discontinued-product URLs, and conflicting canonical tags.

Do not assume platform defaults are always wrong. The configuration, theme, app stack, catalog size, and merchandising choices determine the actual risk. Audit the live output rather than relying on generic platform advice.

Sold-out products require a deliberate decision. Temporary stock gaps may justify retaining the page and offering alternatives; permanently discontinued products may need a relevant redirect, a retained informational page, or an intentional 404. The answer depends on demand, backlinks, replacement products, and shopper value.

The 80/20 Rule for eCommerce Technical SEO

The 80/20 rule in SEO is a prioritization principle: a small number of problems often drives a disproportionate share of lost visibility and revenue. In eCommerce, those problems are usually template-level defects on high-value collections and products, not isolated warnings on obscure URLs.

Find the high-leverage fixes by combining affected URL count with impressions, organic sessions, conversions, and template scope. A slow product-gallery script used across 2,000 revenue pages is a stronger candidate than a handful of missing alt attributes on low-traffic images.

Technical SEO Audit FAQs

How Do You Perform a Technical SEO Audit?

Establish baselines, crawl the site, and compare crawler findings with Google Search Console. Then check crawlability and indexability, architecture, duplicates and canonicals, speed and Core Web Vitals, mobile usability, HTTPS, metadata, and schema before prioritizing and validating fixes.

The important step is translating patterns into root causes and developer-ready actions. A tool export finds possibilities; an audit makes decisions.

How to Do an SEO Site Audit?

A general SEO site audit includes technical SEO plus content quality, keyword targeting, backlinks, authority, and conversion opportunities. Start with the technical foundation when important pages are blocked, excluded, slow, duplicated, or poorly linked, then expand into broader growth work.

For eCommerce sites, connect every finding to the product, collection, or editorial templates responsible for organic traffic and revenue.

What Is the 80/20 Rule in SEO?

The 80/20 rule means focusing on the smaller set of actions likely to create the largest outcome. In a technical SEO audit, prioritize blockers and template-level problems affecting valuable pages before spending effort on isolated, low-impact warnings.

Use visibility risk, affected URL count, impressions, conversions, and implementation effort to decide what belongs first.

What Is the Best Tool for Conducting a Technical SEO Audit?

No single tool is sufficient. Google Search Console plus a crawler is the best minimum combination because Search Console shows Google’s indexation and performance signals while a crawler reveals sitewide technical patterns.

Add analytics for commercial context, PageSpeed Insights for performance diagnostics, and manual testing for mobile and conversion-path validation.

If your Shopify or WooCommerce store generates $50k+ MRR and technical issues are limiting organic growth, book a no-obligation deep-dive consultation with SEO.DIGITAL to uncover the highest-impact fixes.

Table of Contents