SEO Dictionary

Crawler (web spider)

A crawler — also called a spider or bot, such as Googlebot — is an automated program that follows links across the web to discover pages and read their content. Crawling is the first step before a page can be indexed and ranked.

A crawler starts from known URLs and sitemaps, requests each page, reads the HTML, and follows the links it finds to discover more pages. What it can reach and render determines what search engines know about your site. Server responses matter here: a healthy page returns an HTTP 200, while errors or blocks stop the crawler from seeing content.

You influence crawling with robots.txt (which paths bots may request), internal linking (how easily pages are discovered), and site speed. On large sites this ties directly into crawl budget — the finite attention Googlebot gives your domain.

Crawling is not the same as indexing. A page can be crawled but not indexed if it's thin, duplicated or blocked by a noindex tag. Getting crawling right is the foundation that everything else in technical SEO builds on.

This is a technical SEO topic I handle for clients: Technical SEO →

Work together

Need help with this?

Tell me about your site and where you're stuck. I'll give you an honest read on what would actually move the needle.

Start a conversation →

Open for projects