Skip to content
TrackPodcasts
technologyMar 12, 202625:07failed

Google crawlers behind the scenes

About this episode

Developers often talk about Googlebot as if it were a single program you could just run as "googlebot.exe", but that is not how Google's crawling actually works. In this episode of Search Off the Record, Martin and Gary from the Search Relations team unpack how Google's crawling infrastructure is really built and operated.​
They cover why "Googlebot" is a misnomer and how it relates to a central crawling software-as-a-service used by many Google products​, how crawl behavior is controlled centrally to avoid overwhelming sites (throttling, handling 503s, and "don't break the internet" safeguards)​ and more!
If you build for the web, work on SEO, or just want a more accurate mental model of how Google crawls pages, this behind‑the‑scenes discussion is for you.

Resources:
​Crawlers → https://developer.google.com/crawling 

Episode transcript → https://goo.gle/sotr107-transcript 

Listen to more Search Off the Record → https://goo.gle/sotr-yt  

Subscribe to Google Search Channel → https://goo.gle/SearchCentral 

Search Off the Record is a podcast series that takes you behind the scenes of Google Search with the Search Relations team.

 #SOTRpodcast #SEO #GoogleSearch

Speakers: Martin Splitt, Gary Illyes

Get every episode summarized

Each time Search Off the Record publishes, we email you a written briefing from the transcript — the topics, who appeared, and any specific claims, with the ad reads skipped.

Email me new episodes

Free for 3 shows. No card needed.

Hosts & guests

No transcript yet

This episode has not been transcribed. Request it and it moves to the front of the queue.

Google crawlers behind the scenes

Search Off the Record

0:00
25:07

More episodes

More from Search Off the Record

View all episodes →