Sooner or later, you find yourself staring at Google Analytics thinking: "These numbers should be better." Traffic is flat, rankings barely move, and you have no idea why. That's exactly where a structured SEO audit comes in.
An SEO audit isn't a one-off exercise. It's a tool you should use regularly. In this post, I'll walk you through what really matters, step by step, and show you how to build your own audit without getting lost in the details.
And because optimizing for Google alone is no longer enough in 2026, we'll also look at how to prepare your content for AI systems like ChatGPT, Perplexity and Google AI Overviews.
WHAT IS AN SEO AUDIT?
An SEO audit is a structured analysis of a website that examines its technical foundation, on-page optimization and content quality for weaknesses and untapped potential. The goal is to find out why a site isn't performing in search the way it should, and to turn that into concrete next steps.
A good audit covers five areas:
- Crawlability and indexing (can Google find the page at all?)
- Technical foundation (is the site fast and error-free?)
- On-page optimization (is the content set up for search engines?)
- Content quality (does the page deserve to rank well?)
- Authority and links (is the site seen as credible on the web?)
The order isn't random. What good is great content if Google can't even crawl the page?
┌─────────────────────────────────────────────────────┐
│ SEO AUDIT 2026 │
│ ANALYSIS ORDER │
└──────────────────────┬──────────────────────────────┘
│
┌────────────────┼────────────────┐
▼ ▼ ▼
┌───────────┐ ┌───────────┐ ┌───────────────┐
│ CRAWL- │ │ INDEXING │ │ TECHNICAL │
│ ABILITY │ │ │ │ FOUNDATION │
│ │ │ │ │ │
│ robots.txt│ │ Canonical │ │ Core Web │
│ Sitemap │ │ Noindex │ │ Vitals │
│ Structure │ │ Redirects │ │ HTTPS/Mobile │
└─────┬─────┘ └─────┬─────┘ └──────┬────────┘
│ │ │
└───────────────┼────────────────┘
▼
┌────────────────────────────────┐
│ ON-PAGE SEO │
│ │
│ Title Tags · Meta Desc │
│ Headings · Images │
│ Internal linking │
└───────────────┬────────────────┘
▼
┌────────────────────────────────┐
│ CONTENT & AUTHORITY │
│ │
│ E-E-A-T · Thin Content │
│ Keyword mapping · Backlinks │
│ Schema Markup │
└───────────────┬────────────────┘
▼
┌────────────────────────────────┐
│ GEO READINESS │
│ │
│ Definitions · FAQ │
│ Quotable statements │
│ Complete answers │
└────────────────────────────────┘
STEP 1: CAN SEARCH ENGINES FIND YOUR WEBSITE?
What is crawlability and why does it matter?
Crawlability describes how easily search engine bots like Googlebot can access a website and discover its pages. If a page can't be crawled, it can't be indexed, and then it won't rank anywhere.
The three most important checkpoints:
robots.txt
This file (at yourdomain.com/robots.txt) tells search engines which areas they shouldn't crawl. The problem: sometimes important pages get blocked by accident. This often happens after a relaunch, or when someone changes the wrong setting in the CMS.
What to check:
- Are important pages allowed?
- Does it reference your XML sitemap?
- Are any important directories blocked (e.g.
/wp-content/or product categories)?
XML sitemap
An XML sitemap is a structured list of all the important pages on your website that you provide to search engines. It helps Google discover all of your content.
Sitemap checklist:
- Does it exist, and is it accessible at
/sitemap.xmlor/sitemap_index.xml? - Has it been submitted in Google Search Console?
- Does it only contain pages that should be indexed (no redirects, no noindex URLs)?
- Is it updated regularly?
Site architecture
Neither users nor bots want to dig around forever to find content. The rule of thumb: every important page should be reachable within three clicks from the homepage. Orphan pages (pages with no incoming internal links) often aren't found by Google at all.
STEP 2: CHECK INDEXING, AN OFTEN UNDERESTIMATED STEP
Crawlability and indexing often get mixed up. A page can be crawled and still not make it into Google's index. It works the other way around, too: a page blocked in robots.txt can still end up in the index if other pages link to it.
Quick test: Search for site:yourdomain.com on Google. The number of results gives you a rough idea of how many of your pages are indexed. Does it roughly match the number of pages you expect?
Typical indexing problems:
| Problem | Possible cause |
|---|---|
| Important pages not indexed | Noindex tag set by accident |
| Too many pages indexed | Parameter and filter pages without canonicals |
| Duplicate content in the index | Missing or incorrect canonical tags |
| Soft 404s | Empty pages return HTTP 200 instead of 404 |
Canonical tags deserve special attention. A canonical tag is an HTML element that tells search engines which version of a page is the "original" when similar or identical content is available under different URLs. For example, if a product is available under several URLs (with and without a trailing slash, with and without parameters), the canonical makes sure Google picks the right one.
STEP 3: TECHNICAL SEO, THE FOUNDATION FOR EVERYTHING ELSE
Core Web Vitals: Google's benchmark for user experience
Core Web Vitals have been an official Google ranking factor since 2021. The three metrics:
LCP (Largest Contentful Paint): How long does it take for the largest visible element on the page to load?
- Target: under 2.5 seconds
- Common causes of poor scores: uncompressed images, no CDN, slow servers
INP (Interaction to Next Paint): How quickly does the page respond to user input?
- Target: under 200 milliseconds
- Common causes: too much JavaScript, blocking scripts
CLS (Cumulative Layout Shift): How much does the layout shift while the page loads?
- Target: under 0.1
- Common causes: images without fixed dimensions, ad banners that load in late
Where to measure? Google PageSpeed Insights is the easiest place to start. For deeper analysis, WebPageTest is worth a look.
In my detailed Core Web Vitals guide, I show you step by step how to break down and fix each of the three metrics.
Other technical factors that often get overlooked
HTTPS and SSL: A basic requirement these days. Still, some sites have invalid certificates or mixed content (resources loaded over HTTP even though the page itself uses HTTPS). Mixed content warnings in the browser are a bad sign.
Mobile-friendliness: Google primarily evaluates websites based on their mobile version (mobile-first indexing). In other words: if your site shows less content on a smartphone than on a desktop, Google sees less, too.
URL structure: Good URLs are short, descriptive and include the relevant keyword. A bad URL looks like this: yourdomain.com/?p=1234. A good one looks like this: yourdomain.com/seo-audit-checklist.
STEP 4: ON-PAGE SEO, THE PART MOST PEOPLE KNOW
Title tags: your first impression in Google
The title tag is what appears as the clickable blue headline in the search results. It's one of the most important on-page SEO factors there is.
What matters:
- Every page needs a unique title tag
- The primary keyword should come as early as possible
- The ideal length is 50 to 60 characters (otherwise the title gets cut off in the search results)
- The title should invite clicks, not just be optimized for bots
- The brand name usually goes at the end
Common mistakes:
Duplicate title tags are a classic problem, especially on larger websites. If five pages share the same title, Google can't tell which one is relevant for a query.
Keyword stuffing (cramming in the same keywords over and over) doesn't work. Google is smart enough to recognize it, and it just reads awkwardly to users.
Meta descriptions: not a ranking factor, but important for clicks
The meta description isn't a direct ranking factor. But it does influence whether someone clicks on your result or scrolls past it. A good meta description:
- Is unique for each page
- Is 150 to 160 characters long
- Contains the primary keyword (Google often highlights matching terms in bold)
- Makes a clear promise or includes a call to action
If you don't write a meta description, Google pulls a snippet from your page content. Sometimes that works, but usually it doesn't.
Heading structure: using H1 to H6 correctly
Every page needs exactly one H1 (the main heading). It should contain the primary keyword and make it clear what the page is about. Below that come H2s for the main sections, H3s for subsections, and so on.
What I see again and again in audits:
- Multiple H1 tags on one page (often because both the page title and the logo are marked up as H1)
- Headings that don't describe the content below them and are only used for visual styling
- Skipped levels in the hierarchy (jumping from H1 straight to H3)
STEP 5: IMAGES AND INTERNAL LINKING
Optimizing images
Images are often the biggest drag on page speed. But they're also an opportunity to get traffic from Google Images.
Image checklist:
- Every image has alt text (a short description of the image for screen readers and search engines)
- File names are descriptive (
seo-audit-screenshot.jpg) rather than cryptic (IMG_4832.jpg) - Images are compressed (tools: Squoosh, TinyPNG)
- A modern format is used where possible (WebP is smaller than JPEG at the same quality)
- Lazy loading is enabled (images outside the visible area only load when they're needed)
Internal linking: the underrated SEO tool
Internal links (links from one page on your website to another) do two things: they help users find more relevant content, and they show Google which pages matter and how they relate to each other.
Typical problems:
Orphan pages: Pages that no internal link points to. Google may never find them.
Anchor text: The anchor text (the clickable text of a link) should describe where the link goes. "Click here" tells Google nothing. "SEO audit checklist" does.
Over-optimization: If every internal link to a page uses exactly the same keyword-heavy anchor text, it looks unnatural.
I'll audit your entire website and give you every fix, prioritized by impact and effort, so your team can get started right away.
REQUEST AN SEO AUDITSTEP 6: CONTENT QUALITY, THE HARDEST PART
What is E-E-A-T and why does it matter?
E-E-A-T stands for Experience, Expertise, Authoritativeness and Trustworthiness. It's Google's framework for assessing whether content is trustworthy and backed by real expertise. The concept comes from Google's Search Quality Rater Guidelines, the handbook human quality raters use to evaluate websites.
The four dimensions at a glance:
| Dimension | What Google looks at | How you demonstrate it |
|---|---|---|
| Experience | Does the author have first-hand experience? | Your own examples, screenshots, personal insights |
| Expertise | Is the author qualified? | An author page with credentials, depth in the content |
| Authoritativeness | Do others cite the site as a source? | Backlinks from relevant sites, mentions |
| Trustworthiness | Can the site be trusted? | HTTPS, legal notice, privacy policy, accurate information |
What is thin content and why does it hurt rankings?
Thin content means pages with little unique or useful information. Classic examples: WordPress tag pages that just list post titles, product pages that copy the manufacturer's description, or pages that only exist to target a keyword without offering any real value.
Google wants to rank content that actually answers the search query. If your page doesn't, sooner or later it'll be overtaken by pages that do it better.
Keyword cannibalization: when you get in your own way
Keyword cannibalization happens when several pages on your website target the same keyword. Google can't decide which page to rank, and they all end up doing worse than a single strong page would.
The fix: a keyword map where every page has one clear primary keyword and one clear job.
STEP 7: SCHEMA MARKUP, AN OFTEN MISUNDERSTOOD TOPIC
Schema markup (also called structured data) is code you add to your website so Google can better understand your content and display it as rich results (enhanced search results). Examples include star ratings, product details or recipe cards right in the search results.
An important note for anyone who runs audits or commissions them:
Some websites only add schema markup via JavaScript, for example through Google Tag Manager, certain plugins or JavaScript frameworks. That means if you check the page source with a simple tool, you often won't see the markup at all, even though it's there. An audit that reports "no schema found" on that basis is wrong.
How to check schema markup properly:
- Google's Rich Results Test (renders JavaScript)
- Screaming Frog with JavaScript rendering enabled
- Directly in your browser: open the developer tools, go to the "Elements" panel and search for
application/ld+json
STEP 8: COMMON PROBLEMS BY PAGE TYPE
Not every SEO problem applies to every type of site. Here's a quick overview:
SaaS and product pages
- Product pages often have too little content (thin content)
- The blog doesn't link to the product pages
- There are no comparison pages (e.g. "Tool X vs. Tool Y") or alternatives pages
E-commerce
- Category pages have no unique text, just product listings
- Product descriptions are copied word for word from the manufacturer
- Faceted navigation (filters) creates thousands of duplicate URLs without canonicals
Content sites and blogs
- Old posts never get updated
- Similar articles cannibalize each other's keywords
- No topic clusters, so there's no clear topical focus
Local businesses
- NAP data (name, address, phone number) is inconsistent across platforms
- No LocalBusiness schema markup
- No location-specific landing pages
TOOLS FOR YOUR SEO AUDIT: WHAT YOU REALLY NEED
Free:
- Google Search Console: A must. Shows indexing issues, Core Web Vitals, top keywords and crawl errors.
- Google PageSpeed Insights: Core Web Vitals plus suggestions for improving page speed.
- Rich Results Test: For checking schema markup.
- Bing Webmaster Tools: Often forgotten, but gives you additional data.
Paid (worth it once your site reaches a certain size):
- Screaming Frog: Crawls your entire website and gives you all the technical data. Essential for larger projects.
- Ahrefs, SISTRIX or Semrush: For backlink analysis, keyword research and competitor comparisons.
- Sitebulb: Similar to Screaming Frog, with better reporting.
SEO AND GEO: WHY YOU SHOULD ALSO OPTIMIZE FOR AI SYSTEMS IN 2026
This is the part that's still new to a lot of people.
Classic SEO optimizes for Google. GEO (Generative Engine Optimization) means preparing your content so that AI systems like ChatGPT, Perplexity or Google AI Overviews understand it, cite it and recommend it as an answer.
In practice, that means:
Clear definitions: Explain technical terms right in the text. AI systems look for precise definitions they can use to answer user questions.
Questions and answers: Phrase headings as real user questions and answer them directly below. Not three paragraphs later.
Quotable statements: "AI is changing a lot" isn't quotable. "According to a 2025 Pew Research Center study, Google users who saw an AI summary clicked on a regular search result in only 8% of visits, compared with 15% without one" is.
Complete answers: A post should cover its topic thoroughly enough that an AI system sees it as the best available answer. That means providing context, answering sub-questions and not leaving out counterarguments.
FAQ section: An FAQ section at the end can improve your chances of showing up in AI-generated answers.
PRIORITIZATION: WHERE TO START?
You can't fix everything at once. Here's a practical order:
Fix immediately (blocking rankings):
- Crawl errors and robots.txt blocks
- Missing or incorrect canonicals
- Noindex tags on important pages
- Redirect chains and loops
High priority (noticeable impact):
- Improve Core Web Vitals
- Optimize title tags and H1s
- Expand or merge thin content
- Strengthen internal linking
Quick wins (fast to do, immediate effect):
- Add missing alt text
- Write meta descriptions where they're missing
- Submit your sitemap in Search Console
Long-term (ongoing work):
- Strengthen E-E-A-T signals
- Build topic clusters
- Build backlinks systematically
CONCLUSION: AN SEO AUDIT ISN'T A ONE-TIME PROJECT
What strikes me most about SEO audits: they almost always uncover problems nobody had on their radar. A robots.txt that accidentally blocks important pages. A canonical pointing the wrong way. Ten pages targeting the same keyword and pushing each other out of the rankings.
But the real value isn't just in finding these problems. It's in prioritizing them. What's blocking your rankings right now? What has the biggest impact in the medium term? What can wait?
If you're doing your first audit, start with Google Search Console. It's free, it comes straight from Google, and it shows you more than most people realize. From there, work your way through the areas I've covered here, step by step.
And if you want to be visible in AI systems too: write comprehensively, define terms clearly, and answer questions as if you were explaining them to a smart friend. That's good SEO and good GEO at the same time.