WANT TO GROW YOUR ORGANIC SEARCH CHANNEL? SCHEDULE YOUR DEMO NOW.

HOW TO USE GOOGLE'S CRAWL STATS REPORT FOR TECHNICAL SEO

WHAT IS THE GOOGLE CRAWL STATS REPORT?

The new Google Crawl Stats Report — released in November 2020 — helps developers, webmasters, and SEOs understand Google’s crawling experience on their websites. It is a major upgrade over the previous report, which only showed three metrics in isolation: pages crawled per day, data downloaded per day, and time spent downloading pages per day.

Google lighthouse crawl requests

The new version looks more like the rest of Google Search Console and there is a lot of critical information for technical SEO here. The report provides in-depth data on how Google crawls your site, how your site reacts to any given crawl, and its overall technical health of your site — as well as advice on how to debug any associated issues.

WHERE DO I FIND THE CRAWL STATS REPORT?

To access the new Google Crawl Stats report, you need Search Console and access to the property. Click into the property. On the left-hand rail under property settings, click on “crawl stats” to open the report.

You will see a window for crawl history over the last 30 days which reports on metrics like total crawl requests, total download size, and average response time for requested URLs. There will be another window for host status issues associated with your robots.txt file, DNS resolution, and server connectivity. And you will also see a window for crawl breakdown by the response, file type, purpose, and agent type.

WHAT DOES THE SEARCH CONSOLE CRAWL STATS REPORT MEASURE?

The crawl stats report measures how bots interact with your site. At the top of this report — above the fold — you have a graph and 3 metrics shown: total crawl requests, total download size, and time spent downloading a page.

Here is a break down of the various crawl stats metrics:

TOTAL CRAWL REQUESTS

Total crawl requests measures the total number of crawls for URLs on your site in a given time period to indicate how frequently Google is crawling your site. These include both successful and unsuccessful crawl requests. An unsuccessful request might result from DNS issues, server connectivity issues, redirect loop issues, or fetches never made because of an unavailable robots.txt file.

TOTAL DOWNLOAD SIZE

Total dowload size is a metric that reflects how much content Google is downloading during its crawling process in the given time period. If you have high averages, Google is crawling your site often and downloading a lot of content. However, those high averages could also mean Google is taking too long to crawl your site. That said, good average response times, would offset this issue as they are a good indication of a site that is efficient to crawl.

AVERAGE RESPONSE TIME

Average response time tells you how long it takes for a search engine to request page content. The lower the number, the less time Google is spending on your site and thus, crawling and indexing at a faster rate. To improve this number, there are a number of optimizations that can be made. For example, you can block unnecessary pages from being crawled by modifying your robots.txt file, or cutting bloated content or code.

HOST STATUS

There is also a host status section that shows whether Google encountered any availability issues with your robots.txt file, DNS resolution, or server connectivity while attempting to crawl the site in the last 90 days. For each of these points, it will indicate whether your site has had an “acceptable fail rate,” an “acceptable fail rate recently, but high in the past,” or “a high fail-rate in the last week.” You can inspect the response tables to see what the specific problems are — if there are any.

SOLVING HOST STATUS ISSUES

Here are some courses of action you can take for robots.txt files, DNS resolutions and server connectivity:

CRAWL REQUESTS BREAKDOWN

Below the fold, there is a crawl requests breakdown by response, file type, purpose, and Googlebot type. Click into any of these grouped Data Type Entries to see a list of example URLs for that type.

CRAWL RESPONSES

Crawl responses show the responses that Google receives when crawling your site. They are grouped together by code (like 200, 301, 302, 404, and 5xx) and given a percentage to represent how much of the crawl budget was used on them. It is critical to determine here what percentage of your crawl budget is used on non-200 responses and to make changes accordingly

FILE TYPE

File type shows the percentages of crawl budget used on various types of files like HTML, Javascript, CSS, image, video, and audio to name a few. Understanding how frequently Google requests specific types of resources like Javascript and CSS can better inform a number of different technical SEO strategies, like the type of rendering you want to use on your site.

crawl stats by file type

CRAWL PURPOSE

Crawl purpose indicates whether Google is requesting a URL they have never crawled before (discovering new content) or returning to a known page looking for refreshed content. If you have recently added a lot of new content or submitted a new sitemap, you will likely see an increase in “discovery” crawls in this breakdown. If you have pages with rapidly changing content, you will likely see larger percentages of “refresh” crawls in this breakdown. Understanding the purpose behind a crawl and seeing example URLs included can help you drill down into what pages are receiving priority and whether you need to fix any issues in your sitemap, robots.txt file, or internal linking system to help Google access important content.

crawl stats by purpose

GOOGLEBOT TYPE

Googlebot Type shows which types of crawlers — mobile, desktop, image, video, page resource load, adbot, storebot, etc — are accessing your site and how often they are doing so. It depends on the site, but the majority of crawls will likely come from the mobile or desktop bot to simulate user experience on those devices.

google-bot-type-crawl-stats

HOW CAN I USE THE GOOGLE CRAWL STATS REPORT FOR SEO?

This Google Search Console Crawl Stats report is a big help for technical SEO, which of course deals heavily with crawling and indexing of the website. If Google can’t properly crawl your site, they will not be able to index new pages or detect changes to old pages and consider the content for ranking purposes. So this new report gives actionable data to use when debugging crawling and general site performance issues. The report also makes it way easier now to diagnose hosting problems, resources eating up too much crawl budget, 404 errors, and the like. You are more clearly seeing your website from Google’s point of view. With this data on how they crawl our site and how our site responds, we can make more effective, more informed optimizations.

HERE ARE SOME EXAMPLE USE CASES:

In general, it is also an accessible and easy report to read — particularly if you don’t have a technical background or access to your log files. Every section breaks down — sometimes even in color-coding — what is working well, what needs to be addressed, and why with best practices and tips included.

CONCLUSION: CRAWL STATS AND GENERATIVE ENGINE OPTIMIZATION

The Google Crawl Stats report is now a foundational tool for Generative Engine Optimization (GEO). Its value extends beyond traditional debugging to signal site quality to AI crawlers like GPTBot and Google-Extended. A high percentage of non-200 status codes, such as 301 redirects or 404 errors, directly impacts your crawl budget. This inefficiency tells generative AI that your site is poorly maintained and not a reliable source. As noted in a 2024 Search Engine Land analysis, monitoring Googlebot activity via the Crawl Stats report is key to identifying these actionable insights. For large, enterprise websites, this data is not just diagnostic; it is predictive of your visibility in AI-driven search. A clean, efficient crawl is a prerequisite for becoming a citable authority for Large Language Models. Optimizing these technical signals ensures that when AI engines look for definitive answers, they can easily find and trust your content.

How does this report stack up to traditional log file analysis? The big differentiator is that you can now determine the purpose of each Google visit to your site. That’s not possible to glean from log files. And as we have already covered, this report is a massive upgrade over the old one. But there are shortcomings. First and foremost, this report still only records Google’s activity on your site. Furthermore, you are just receiving a sample of crawled URLs — not the full list. And then there are other issues — like the fact that you can’t toggle the date range for historical numbers, drill down on geographical regions, or access information via API yet. So log file analysis is still important. Log files record every request to your site, like crawling activity from Bing for example. And the data is also more accurate in the moment — sometimes 20-40% more accurate, as Google’s data can lag up to a week. So consider this report a valuable, but still limited, tool in your technical SEO kit.

For information on how to automatically action technical recommendations from this crawl stats report, check out our software platform, Huckabuy Cloud.

  • www.womenintechseo.com
    We all meet at the same crossroads. Suddenly, dozens, hundreds, or millions of URLs are stuck in the dreaded Discovered, not crawled, or Crawled not indexed death zone in Google Search Console (GSC).
    Navigating the "Death Zone" with Crawl Stats
    Tomek Rudzki's April 16, 2021 article, "It’s not you - it’s Google Search Console | Women in Tech SEO," truly resonates with the collective...
  • Discourse Meta
    There could be soooooo many reason for this.Is the googlebot actually crawls your site ? check mysite.com/admin/reports/web_crawlersIs the googlebot blocked or rate limited?
    Connecting the Dots: Using Crawl Stats to Diagnose Indexing Issues
    The article "Why isn't Google Indexing Discourse? SEO concerns" raises critical questions about Googlebot's interaction with a site, such as whether Googlebot is actually...
  • Siteguru
    By Rick van Haasteren5 min read - last updated 2 Mar 2023Fix your crawl, and everything else becomes a breeze.
    Insights on Google Search Console Crawl Stats
    We appreciate Rick van Haasteren's concise 5-minute read, "Google Search Console Crawl Stats Report," last updated on March 2, 2023. His article rightly highlights...
  • Sitebulb
    Tech SEO Crawling Aishat AbdulfatahPublished May 19, 2025Massive thanks to Aishat Abdulfatah for this guide on crawl budget optimization. Learn 8 expert-proven strategies for optimizing your crawl budget.You cannot simply force a Google crawl.
    Enhancing Crawl Optimization with Expert Strategies
    It’s incredibly valuable to see how closely our exploration of Google’s Crawl Stats Report aligns with the latest expert thinking in the industry. We...
  • LION Publishers
    This guest post was written by Daniel Petty, the director of audience strategy for ProPublica.
    Elevating E-E-A-T with Google Crawl Stats for Publishers
    We appreciate Daniel Petty's insightful contribution in his article, "7 Technical SEO Essentials for Local Publishers," where he astutely highlights the indispensable role of...
  • www.seoforgooglenews.com
    Barry AdamsOct 04, 20231763ShareThis is an expansion of a concept I touched upon in my previous newsletter, but I felt it needed a proper deep-dive to explain its intricacies.I have to start with a disclaimer: Much of what follows is speculation, based on my experiences and those
    Enhancing Your SEO Strategy: Pre-Publication Optimization Meets Crawl Stats
    We were particularly interested to see Barry Adams’ piece from October 4, 2023, discussing the critical importance of optimizing articles *before* they are even...
  • LinkedIn
    Tim Soulo8moReport this postHow does Ahrefs' marketing team use Ahrefs? Here's a good example: I noticed that our blog has 4700+ pages that were crawled by ahrefs_bot, but only ~1600 of them were seen ranking in Google. How so?
    Unlocking Ranking Potential: Addressing Crawl-to-Rank Discrepancies with Crawl Stats
    The Ahrefs team's discovery of 4700+ crawled pages with only ~1600 ranking is a classic example of a crawl-to-index challenge, leading to what Patrick...
  • LinkedIn
    Chris Long2yReport this postTechnical SEO Tip: You can use Search Console's hidden "Crawl Stats" report to get an estimate of your site's allocated crawl budget: Crawl budget is the concept that search engines will only allocate so many crawling resources to your site.
    Chris Long on Leveraging Crawl Stats for Crawl Budget
    We appreciate Chris Long's insightful tip from two years ago, highlighting the value of Google Search Console's "Crawl Stats" report for understanding and estimating...
  • SUSO Digital
    Author: Deep Shah, Senior Project Manager Ensuring that your website ranks well in Google Search involves more than just quality content and effective SEO strategies.
    Integrating Crawl Stats with Essential Technical SEO Practices
    As we delve into the intricacies of Google's Crawl Stats Report, it's clear that understanding how Googlebot interacts with your site is crucial for...
  • MarketingSyrup
    Some people might call me crazy.I call myself curious.That’s why I did a thing:I disabled the crawling of my main website.
    Insights from a Google Crawling Experiment
    As we delve into the intricacies of Google's Crawl Stats Report, it's fascinating to see how real-world experiments, like the one detailed in the...
  • Search Engine Journal
    Crawl stats is a powerful report within Google Search Console, though it's not accessible from within the main interface. Here's why you need to find it.Tomek RudzkiApril 16, 2021⋅7 min readTomek Rudzki Head of R&D at OnelyBio Follow652SHARES15KREADSThere is one report in Goo
    A Comparative Analysis of Google's Crawl Stats Insights
    Having read the article "5 Top Crawl Stats Insights in Google Search Console", I found it to be a useful resource for understanding the...
  • SEOSLY
    Updated: June 9, 2023.Here I’m showing you step by step how to audit your site with Google Search Console only.SEOs when doing SEO audits always use a site crawler like Sitebulb or/and Semrush Site Audit as the main tool and a bunch of other SEO tools.But is it possible to perfor
    Leveraging Google's Crawl Stats Report for a Comprehensive SEO Audit
    Having read the article on "How To Audit A Site With Google Search Console Only - SEOSLY", I appreciate the comprehensive approach it takes...