What this checklist covers (and what it skips)
Technical SEO is the set of site engineering choices that let search engines discover, understand, and fairly evaluate your pages. For B2B and SaaS marketing sites, that usually means a Next.js or similar app with blogs, docs, pricing, and product pages—not a thousand doorway locations. This checklist prioritises durable hygiene over tricks.
It does not replace content strategy, link earning, or product-led growth. Thin pages with perfect Lighthouse scores still struggle. Conversely, excellent articles trapped behind broken canonicals or noindexed templates waste budget. Use technical work to remove friction from content that deserves to rank.
Work the list in layers: host and HTTPS first, then crawl/index rules, then templates (title/canonical/hreflang), then performance, then structured data. Fixing schema while the site returns soft-404s for filtered URLs is rearranging deck chairs.
Document owners. Marketing often owns copy; engineering owns robots, headers, and deployments; design owns CLS. Without named owners, checklist items reopen every quarter. Treat this as a living release gate, not a one-off PDF.
Crawl, indexation, and URL health
Confirm a single preferred host (apex or www) with HTTPS and consistent redirects. Mixed hosts create duplicate signals and dilute analytics. Check that HTTP → HTTPS and non-preferred host → preferred host use stable permanent redirects. Avoid redirect chains longer than necessary.
Audit robots.txt: allow important marketing paths; disallow only what you truly want uncrawlable (for example internal search result URLs or staging leftovers). Do not block CSS/JS required to render content if you rely on modern rendering. Keep a human-readable comment block so future deploys do not “clean up” critical rules accidentally.
XML sitemaps should list canonical indexable URLs only—no redirects, no noindexed URLs, no parameter junk. Split large sets if needed; reference sitemaps from robots.txt. Monitor coverage in Search Console or Bing Webmaster for sharp drops after releases.
Canonical tags must match the URL you want indexed. Self-canonicals on templates are fine; accidental cross-canonicals from shared layouts are not. Pagination, faceted filters, and UTM-heavy links need a clear policy: usually canonical to the clean URL and avoid indexing infinite filter combinations.
Status codes tell the truth: soft 404s that return 200 with “not found” copy confuse crawlers. Auth walls, geo blocks, and bot challenges should not silently empty important pages for Googlebot if those pages are meant to rank. Log bot fetches during releases when you change edge rules.
Internal linking should reach money pages from crawlable HTML—not only from client-side menus that never appear in the initial document. Footer and contextual links help; orphan landing pages do not. After IA changes, crawl for broken links and update hubs.
Rendering, JavaScript, and performance
Prefer server-rendered or statically generated HTML for primary marketing content. Heavy client-only pages risk delayed indexing or incomplete snapshots. If a page needs personalisation, keep the indexable core in the initial HTML and hydrate enhancements afterward.
Core Web Vitals (LCP, INP, CLS) are user metrics that also correlate with healthy pages. Optimise LCP candidates (hero image or headline block), reduce long tasks that hurt INP, and reserve space for images/embeds to limit CLS. Measure on mobile field data when available; lab scores are directional only.
Images: modern formats, sensible dimensions, priority hints for the LCP image, and lazy-loading for below-the-fold media. Do not lazy-load the LCP image. SVGs and icon fonts should not cause layout thrash. Compress marketing screenshots; they are a common LCP culprit on SaaS sites.
Third-party scripts (analytics, chat, A/B tools) often dominate main-thread time. Load non-essential tags after consent and interaction where policy allows; audit tag managers quarterly. A “temporary” pixel that stays forever is still a performance regression.
Caching and CDN: correct cache headers for static assets with hashed filenames; careful caching for HTML that changes per deploy. Stale HTML with new asset hashes causes broken CSS—coordinate cache purge with release. Prefetch sparingly; speculative prefetch of entire SPAs can waste bandwidth.
On-page technical signals that still matter
Unique title tags and meta descriptions per indexable URL. Titles should be human-readable and stable—not keyword soup. Descriptions influence CTR more than rankings; write them like honest ads for the page. Avoid identical titles across hundreds of programmatic pages without differentiating tokens that remain truthful.
Heading hierarchy: one clear H1 that matches the page promise; logical H2/H3 structure. Do not hide the only H1 in a client component that fails to render. For docs and blogs, predictable heading structure also helps AI citation systems parse sections—related to GEO work but useful for classic SEO too.
Canonical language and hreflang if you truly maintain locale variants. Incorrect hreflang is worse than none. If you only have English, do not invent locale clusters. For regional marketing pages, ensure content differs meaningfully—not only a city name swap—if you want them indexed as separate URLs.
Open Graph and Twitter cards do not replace SEO tags but affect sharing previews that drive referral traffic. Keep og:image dimensions consistent; stale social caches after major redesigns may need manual refresh in platform debuggers.
Pagination and infinite scroll: provide crawlable next links or “view all” equivalents for content you want indexed. Infinite scroll alone is a poor discovery mechanism for crawlers.
Structured data, entities, and trust pages
JSON-LD for Organization, WebSite (with SearchAction only if you have a working on-site search), Article/BlogPosting for editorial content, FAQPage only when FAQs are visible on the page, Product/SoftwareApplication when accurate for product pages, and BreadcrumbList when breadcrumbs exist in the UI. Invalid or invisible FAQ schema is a trust risk.
Keep entity facts consistent: legal name, logo, sameAs profiles, and contact points should match footer and about pages. Inconsistencies confuse knowledge panels and AI answer engines alike. Update structured data when you rebrand.
Validate with rich-results testing tools and by sampling live HTML—not only the CMS preview. Frameworks sometimes double-inject JSON-LD. One clean graph beats three conflicting ones.
Do not mark up reviews you do not display, prices you do not show, or how-to steps that are not on the page. Aggressive markup without visible content is a common spam pattern and not worth the risk for B2B brands.
SaaS-specific technical traps
App subdomains and logged-in product URLs often should be noindexed. Marketing site and application should have clear robots policies so you do not index empty dashboards or tokenised invite links. Staging and preview deployments must be auth-gated or noindexed—never accidentally left in sitemaps.
Docs and changelog sprawl: versioned docs can create near-duplicates. Canonicalise superseded versions or noindex them intentionally. Changelog pagination should not create thin date URLs without substance.
Programmatic SEO pages (integrations, templates, locations) need unique value, internal links, and indexation controls. If similarity is high, consolidate or improve content before scaling URL count. A similarity gate in CI is healthier than a manual cleanup later.
Pricing pages that personalise heavily should still expose a default indexable state. Gatekeepers that block all bots from pricing can remove a high-intent page from search. If legal requires geo pricing differences, implement carefully with distinct URLs or clear defaults.
International sales sites sometimes ship /us /uk mirrors with thin differences. Prefer substantive localisation or a single global page with clear currency/plan notes over doorway locale farms.
Release ops: how to work this checklist continuously
Add a short SEO smoke checklist to every marketing deploy: robots.txt reachable, sitemap returns 200, homepage and one article render titles/canonicals, no accidental noindex on production, and LCP image still fine on a mid-tier mobile profile. Automate what you can in CI; keep a human pass for IA changes.
Monitor Search Console and Bing Webmaster for coverage anomalies, spike in soft 404s, and manual actions. IndexNow can accelerate discovery of changed URLs on participating engines—it does not replace sitemaps or fix thin content.
When migrating URLs, ship a redirect map, update internal links and sitemaps in the same release, and keep redirects long enough for signals to transfer. Changing URL structures quarterly for “freshness” usually hurts more than it helps.
Illustrative effort note (assumptions stated): a healthy SaaS marketing site can clear foundational crawl/index items in a focused engineering sprint if the stack is already modern; performance and template debt may take longer phased work. Exact timelines depend on CMS complexity, tag sprawl, and how many legacy redirects exist. Prefer measured fixes over a vague “SEO redesign.”
HiMat’s technical SEO and GEO service pairs crawl/index hygiene with entity-aware content structure. Use free tools like the Core Web Vitals checker, robots.txt generator, XML sitemap generator, schema markup generator, and canonical URL generator to spot issues between audits. For a full pass, start at /schedule—and keep this checklist as your team’s definition of done for marketing releases.