Glossary · SEO

Crawl Budget

krawl BUJ-itnoun

Crawl budget is the number of pages a search engine will crawl on your site within a given time frame.

Part of speech
noun
Pronunciation
krawl BUJ-it
Origin
From 'crawl,' how search bots move through links, plus 'budget,' Old French 'bougette' for a purse. It is the crawling a site is allotted.

What is Crawl Budget?

Crawl budget is the number of pages a search engine will crawl on your site within a given time frame. Search engines have finite resources and countless pages to visit across the web, so they allocate a rough limit to how much of any single site they will fetch before moving on. For most small sites this limit is generous enough that it never becomes a concern. For large sites with many thousands or millions of pages, crawl budget becomes a real constraint that determines how quickly new and updated content gets discovered and how much of the site ever enters the index at all.

Crawl budget is shaped by two forces. The first is crawl capacity, the amount of crawling a site can handle without being overloaded; a fast, reliable server that responds quickly invites more crawling, while a slow or error-prone one causes the search engine to back off to avoid straining it. The second is crawl demand, how much the search engine actually wants to crawl a site, which rises with the site's popularity, freshness, and the perceived value of its pages. The practical result is a limited allotment of crawling that must be spent across your entire site. If that allotment is wasted on low-value pages, the pages you care about may be crawled infrequently or missed.

The term joins "crawl," the way search bots move from link to link discovering pages, with "budget," from the Old French "bougette," meaning a small purse or bag. The financial metaphor fits: like any budget, it is a finite pool of a scarce resource that has to be spent wisely, and overspending in one area leaves less for another. The concept became a defined SEO concern as large sites grew complex enough to generate far more crawlable addresses than search engines were willing to fetch.

For a business, crawl budget matters most at scale. If a large site squanders its budget crawling duplicate addresses, endless filter combinations, or thin pages, search engines have less capacity left to find genuinely important new content, which delays indexing and can leave valuable pages absent from results. On a big e-commerce catalog or a sprawling publisher, efficient crawling directly affects how fast products and articles appear in search and how completely the site is represented. Managing it well means your best content gets seen promptly.

The nuances separate real concern from misplaced worry. Most sites simply do not have enough pages for crawl budget to be a limiting factor, and obsessing over it wastes effort better spent on content and links. Where it does matter, the goal is efficiency: reduce the number of low-value crawlable addresses through clean URL structure, careful handling of parameters, blocking pointless pages from crawling, and pointing search engines to what matters with an accurate sitemap. A common mistake is confusing crawling with indexing; a page can be crawled without being indexed, and blocking a page from crawling does not guarantee it stays out of the index. Crawl budget relates closely to crawlability, indexing, XML sitemaps, and log-file analysis, the last of which lets you see exactly how search bots are spending their budget on your site. Understood in proportion, it is a large-site optimization, not a universal one.

Why it matters

Crawl budget matters most for large sites, where wasted crawling delays indexing of important pages. Managing it keeps your best content fresh in search.