Doing advanced SEO well means following a diagnostic sequence rather than jumping straight to fixes: start with log-file and crawl data to see what’s actually happening, triage the biggest crawl and indexation leaks, restructure internal links around a clear topical architecture, then layer in entity and structured data work. Skipping the diagnostic phase is the single most common reason advanced SEO projects stall — teams start “optimizing” before they understand what’s broken.
This workflow is the same sequence we run at Salterra on client sites with real crawl and indexation problems. It’s built to work whether you’re auditing a 5,000-page e-commerce catalog or a 500-page SaaS site that’s plateaued despite decent content.
The order matters because each step generates the data the next step depends on. Restructuring internal links before you know which pages are actually being crawled means guessing at priorities. Writing new content before you’ve pruned duplicates means burying your best work under material that’s already competing with it. Treat this as a sequence, not a menu to pick from.
Before touching a single page, get at least 30 days of raw server logs — ideally 90. Filter for verified search engine bots (Googlebot, Bingbot, and now GPTBot, ClaudeBot, and PerplexityBot if you care about AI visibility) using reverse DNS verification, not just user-agent strings, since user agents can be spoofed.
Load the filtered logs into a log analyzer — Screaming Frog Log File Analyser is the accessible starting point, JetOctopus or Botify if you’re working at enterprise scale. You’re looking for four things: which URLs get crawled most and least, what percentage of crawl requests hit non-200 status codes, how crawl frequency correlates with your priority page list, and whether crawl activity is trending up or down over the window.
Export the Coverage and Crawl Stats reports from Google Search Console, then run a full crawl of the site with Screaming Frog or Sitebulb in JavaScript rendering mode. Compare three data sets side by side: what your logs say Google actually crawled, what Search Console says Google indexed or excluded, and what a full crawl says exists on the site.
The gaps between these three data sets are where the real problems live. Pages that exist and are linked internally but never appear in logs are a crawl discovery problem. Pages that get crawled repeatedly but never get indexed are a quality or duplication problem. Pages indexed but absent from your internal link graph are orphans surviving on backlinks alone, which is fragile.
With the gap analysis in hand, rank crawl budget leaks by severity. Typical culprits, roughly in order of how often they show up in an audit:
Fix these in order of crawl-budget impact, not ease of implementation. A robots.txt disallow on a parameter pattern might take ten minutes and reclaim more crawl budget than a week of content work.
Document each fix with a before-and-after crawl budget estimate, even a rough one. This matters for two reasons: it builds a case for continued investment in technical work when stakeholders ask what SEO has been doing, and it creates a record you can check against future audits to confirm a fix actually held rather than quietly regressing after the next site update.
Once crawl leaks are under control, turn to the internal link graph. Export it from your crawler and identify orphan pages, pages more than four clicks from the homepage, and any topic clusters that aren’t cross-linked. Then design — deliberately, not organically — a hub-and-spoke structure where each core topic has a pillar page linking out to supporting articles, and every supporting article links back to the pillar and sideways to related spokes.
Rank your priority pages (the ones with the best conversion or strategic value) and audit how many internal links they currently receive relative to less important pages. It’s common to find a low-value blog post outranking your core service page in internal link count simply because it’s older and got linked from more places over time. Correcting that imbalance is often the fastest ranking lever available, because it requires no new content — just rewiring existing pages.
With architecture fixed, shift to content strategy. List the entities — concepts, sub-topics, named tools, related questions — that a genuinely comprehensive resource on your subject would need to cover, and map them against what you currently have. Gaps become your content roadmap; overlaps and near-duplicates become candidates for consolidation.
Reinforce entity signals with structured data (Organization, Person, and topic-specific schema types) and internal linking that explicitly connects related entities to each other, rather than treating every page as an island.
Compare raw HTML source against the rendered DOM for your key templates. Use the URL Inspection Tool in Search Console to check what Google actually sees, and confirm that critical text, links, and structured data are present without requiring JavaScript execution. If they’re not, work with engineering on server-side or hybrid rendering before investing further in content for that template — otherwise new content risks being invisible to crawlers that don’t render JS.
Score existing content on traffic, backlinks, engagement, and topical relevance. Consolidate thin or overlapping pages into stronger single URLs with 301 redirects, and noindex-then-remove anything with no traffic, no links, and no strategic reason to exist. Only after this cleanup should you greenlight net-new content, prioritized by the entity gaps identified in Step Five.
Advanced SEO isn’t a one-time project. Set a recurring cadence — monthly for active sites, quarterly for stable ones — to re-pull logs, re-crawl the site, and re-check Search Console coverage. Track crawl efficiency (percentage of crawl budget spent on indexable, priority URLs) as a standing metric, the same way you’d track rankings or traffic.
Crawl and indexation fixes often show measurable change in Search Console within two to six weeks. Ranking and traffic impact from architecture and content work typically takes two to four months, depending on site size and competitiveness.
The diagnostic steps (one through three) should always come first, since they tell you where the actual problems are. Steps four through seven can be reordered based on what the audit reveals is most urgent for your specific site.
Server log access, Google Search Console, and a crawler with JavaScript rendering (Screaming Frog's free tier covers small sites). Everything else — Botify, OnCrawl, JetOctopus — is an upgrade for scale, not a requirement to start.
Diagnosis and planning, yes. Implementation of rendering fixes, server configuration, and some redirect work usually requires developer involvement, so loop engineering in early rather than after the audit is finished.
Quarterly for most sites, monthly for large e-commerce or publisher sites where inventory and content change constantly and new leaks can appear quickly.
Terry has 30+ years in software and SEO. He’s the founder of Salterra Digital Services and SEO Spring Training, host of the Roundtable SEO Mastermind, and lead instructor at SEO University — teaching the exact tactics his team uses on client work.
This guide is one lesson from the Advanced SEO Techniques course. Get every lesson, framework and checklist — plus the full 38-course catalog — inside SEO University.
Practitioner-focused training across the full digital marketing stack — from technical SEO to conversion optimization and the AI search era. By Salterra Digital Services, since 2011.