Crawl budget describes how much and how often Google crawls a site. It depends on the server's capacity for Googlebot and on Google's demand for individual URLs.
What is crawl budget?
The crawl capacity limit is how many parallel requests the server can safely handle. Crawl demand reflects how much Google wants to recrawl a given URL based on popularity, changes and other signals.
When to care
Crawl budget matters mainly on very large or fast-changing sites. The risk comes from millions of parameter URLs, faceted navigation, internal search and catalogs that change all the time.
On a small business site it's rarely the main reason pages aren't indexed. A block, a wrong canonical, thin content, an orphaned URL or another technical issue is far more common.
When it's a waste of time
If Google crawls your important pages regularly and changes make it into the index, complex crawl budget optimization probably isn't a priority.
How to cut the waste
- Stop endless combinations of filters and parameters from being generated.
- Fix loops, long redirect chains and server errors.
- Keep internal links and sitemaps on canonical URLs.
- Watch server logs and the crawl stats in GSC.
The robots.txt file controls crawling. It doesn't remove anything from the index. A blocked URL can stay indexed if Google knows it from other links.