NEWProduction-ready HTML & CSS templates, live nowBrowse the catalog →
SEO & growth

Technical SEO basics: crawlability, sitemaps and robots.txt

20 February 2026 · 6 min read · By TM Team
TL;DR

Make sure search engines can crawl and index your site: a clean structure, an XML sitemap, a sensible robots.txt, canonical URLs, and no accidental noindex. Technical SEO removes the blockers.

Great content cannot rank if search engines cannot find, crawl and index it. Technical SEO is the unglamorous foundation that makes sure nothing stands between your pages and the search index.

For most small sites it comes down to a handful of things done right.

How search engines process your site

Search engines discover URLs, crawl them, and decide whether to index them. Technical SEO ensures each step works: your pages are linked and listed, crawlers are not blocked, and nothing accidentally tells search engines to ignore your content.

From URL to ranking
Discover
Crawl
Index
Rank

Technical SEO clears the first three steps so content can reach the fourth.

Technical SEO essentials
ItemWhat it does
XML sitemapLists your URLs for search engines
robots.txtGuides what crawlers may access
Canonical tagsPoint to the preferred version of a page
IndexabilityNo accidental noindex on key pages
Clean URLsReadable, stable, logical addresses
HTTPSSecure, trusted connections
Do not accidentally block yourself
The most common technical SEO disaster is a stray noindex tag or a robots.txt rule that blocks your whole site — often left over from development. Check these first when pages will not rank.

A basic technical checklist

  1. 1
    Submit an XML sitemap
    Generate one and submit it in your search console so pages are found.
  2. 2
    Check robots.txt
    Make sure it is not blocking pages you want indexed.
  3. 3
    Verify indexability
    Confirm important pages are not set to noindex.
  4. 4
    Set canonical URLs
    Point duplicate or similar pages to one preferred version.
  5. 5
    Serve over HTTPS
    Secure the site; it is a baseline trust and ranking signal.
  • Content cannot rank if it cannot be crawled and indexed.
  • An XML sitemap lists your URLs for search engines.
  • robots.txt guides crawlers — do not block yourself.
  • Canonical tags resolve duplicate-content confusion.
  • Check for stray noindex tags when pages will not rank.
What is technical SEO?
The work of making a site easy for search engines to crawl and index — sitemaps, robots.txt, canonical tags, indexability and site structure — so content can rank.
Do small sites need to worry about technical SEO?
The basics, yes. A sitemap, a sane robots.txt and no accidental noindex prevent the most common problems. Deep technical work matters more at scale.
Why are my pages not being indexed?
Common causes are a noindex tag, a robots.txt block, or the page not being linked or in the sitemap. Check those first.
technical seocrawlabilitysitemaprobots.txtindexingseo