
Search Off the Record

Loading…

technology
Hosted by Google · technology · EN · 100 episodes
Search Off the Record takes you behind the scenes of Google Search and its inner workings! In each episode, the folks from the Search Relations team will give you background info on the decision-making behind launches, feature prioritization in Search Console, and the projects Google Search teams are working on. They will share fun stories from the many conferences they attend as well as from their day-to-day working life at Google. They will also dive into the currently trending conversations in the SEO community at large. Have a listen!
Required Pod Score for this show. PitchCentric checks your profile against host openness, topical fit, and audience signals before you generate a pitch.
Contact path
Verified email
Booking probability
34%
Guest openness
Selective
Sign up to generate a grounded pitch for Search Off the Record.
Signup to Generate a PitchSearch Off the Record is a technology podcast hosted by Google, with 100 episodes on record and a Required Pod Score of 80. PitchCentric scores this show on Booking Probability, Listen Score, and live audience signals refreshed every 24 hours.
Google hosts Search Off the Record, a technology show with 100 episodes published.
Our AI reads these to draft pitches. Use them as grounding for a pitch that cites a real guest and a specific topic.

Search Off the Record
Search Off the Record
Is your website's internal search feature secretly acting as an open invitation for crawling lots and lots? In this episode of Search off the Record, Martin Splitt and John Mueller pull back the curtain on how internal search results pages can turn into "infinite crawl spaces" that trap Googlebot, waste your crawl budget, and spike your database load. They break down the critical technical differences between blocking search pages via robots.txt versus noindex tags, and expose the massive security liabilities of leaving these pages indexable. In this episode, you'll learn: The "Infinite Crawl Space" Concept: How Googlebot treats internal search functions as infinite crawl spaces that can generate an endless loop of new URLs. Server Strain & Performance: Why uncached internal search pages force constant database lookups and ranking calculations, slowing down your website for real users. Robots.txt vs. Noindex: The distinct technical trade-offs of using a broad robots.txt disallow rule versus a robots meta tag or HTTP header noindex. The Spam Vector Threat: How bad actors search for pharmaceutical, adult, or casino queries on your site to piggyback off your domain authority and display spammy contact info in Google's index. Why 500 Errors aren't great: Why serving a 500 server error code to stop bots on search URLs will backfire and reduce Googlebot's crawl rate across your entire website. Category Pages vs. Search Pages: How systems like Blogger use search parameters for tag landing pages and why you should treat valuable category pages differently. Key Takeaways for SEOs & Developers: Fix Crawling at the Source: Do not use the Google Search Console Removal Tool to handle infinite search URLs; it only filters search results temporarily and does not stop Googlebot from hammering your server. Broaden Your Robots Rules: Use one broad wildcard rule in your robots.txt (like /search?) to cover all query parameters, keeping your file maintainable and clean. Build Real Category Pages: Instead of using internal search parameters as makeshift categories, invest in clean, dedicated category pages to help search engines understand your site's hierarchy. Don't Depend on Auto-Systems: While Google's systems try to automatically recognize and deprioritize infinite spaces, it is slow and unreliable—proactive manual configuration is always safer. Chapters 00:00 - Intro & Greetings 00:45 - Defining Internal Search Results Pages 01:26 - How Googlebot Discovers Search Features & Creates Infinite Spaces 04:15 - Crawl Budget, Server Load, and Database Performance Hurdles 07:32 - Solutions: Robots.txt Disallow vs. Meta Noindex 10:09 - The Fallacy of the Search Console Removal Tool & 404 Pages 12:35 - Why You Should Never Serve 500 Errors to Bots 13:56 - CMS Nuances: Tag Landing Pages and Blog Categories 15:26 - Crafting Broad Robots.txt Patterns and Historical Guidelines 18:17 - When (and When Not) to Allow Indexed Search Pages 20:41 - The Spam Vector Threat: Hacked Content, Casino, & Pharma Exploits 25:06 - Taking Proactive Security Measures for Clients 27:04 - Lazy Search Redirection Hack, Outro & Subscribing Resources Mentioned: Google Search Console (Removal Tool) Do you have a legitimate reason for letting search engines index your internal search results page? Let us know in the comments below, or find us on LinkedIn to share your thoughts! Don't forget to like and subscribe to the podcast on your favorite platform to catch every behind-the-scenes episode from the Search Relations team! Episode transcript → https://goo.gle/sotr113-transcript Listen to more Search Off the Record → https://goo.gle/sotr-yt Subscribe to Google Search Channel → https://goo.gle/SearchCentral Search Off the Record is a podcast series that takes you behind the scenes of Google Search with the Search Relations team. #SOTRpodcast #SEO #GoogleSearch #SearchConsole #SEOTips #CrawlBudget #RobotsTXT #GoogleSearchConsole #WebPerformance #WebSecurity #SearchOffTheRecord

Search Off the Record
Should you panic when your Search Console indexing report is showing pages that aren't indexed? Is a 404 error code always a sign of a broken website? In this episode of Search Off the Record, Martin Splitt and John Mueller from the Google Search Relations team dive deep into the Page Indexing report in Google Search Console. They unpack why treating this report as a static inventory checklist to "fix" things is the wrong response, how to spot massive SEO-ruining hosting or CDN traps, and why a healthy website doesn't actually need a 100% index rate. In this episode, you'll learn: The Indexing Report Shift: Insights from the Search Console team's Hillel on why you should look for trend lines and systemic patterns rather than treating the report as a giant list of errors. When 404s are Good: Why expected 404 errors are technically correct for deleted content, and how to survive the "boss panic" of numbers that won't go down. The Domain Property Advantage: How setting up a domain property handles canonical shifts, www vs. non-www tracking, and performance data much cleaner. The Site Query vs. Search Console: Why the site: query is an artificial tool that might show old domain moves or hreflang swaps for years, making Search Console your only true source of truth. Hosting & CDN Traps: How aggressive bot protections, hidden interstitials, and "Soft 200" error pages completely destroy your crawl data and lead to malicious canonicalization. Discovered vs. Crawled: What it actually means when pages sit in "Discovered/Crawled - currently not indexed," and how to recognize holistic site quality issues over technical bugs. Key Takeaways for SEOs & Developers: Patterns over Inventories: Use the report to verify that your intentional changes (like site migrations or page removals) are processing correctly over time. Forget the Ratio: There is no magic metric for indexed vs. non-indexed pages. Even Google's own developer documentation has a massive chunk of non-indexed content due to intentional choices. Watch Out for Soft Blocks: Ensure your security layers or CDNs aren't serving "Are you a bot?" challenge screens to Googlebot with a 200 success code. Computers Fail (And That's Fine): Minor server blips, failed DNS requests, or temporary 500 errors happen. Google's systems are resilient and will just try again later. Chapters 0:00 - Introduction: The Search Central Live coverage report confusion. 1:45 - Shifting perspectives: Treating Search Console as a pattern tracker, not a checklist. 4:36 - Tracking site migrations and processing data delays. 6:25 - Why 404 errors can be a good thing 8:00 - Handling canonical shifts and the value of Domain Properties. 11:12 - Why the site: query isn't telling you what you think it does (Domain moves and hreflang bugs). 13:38 - CDN bot protection and the absolute nightmare of soft error pages. 17:49 - How to use "marked as fixed". 20:32 - Discovered vs. Crawled Not Indexed: Is it a technical or site quality issue? 25:31 - Debunking the indexed-to-non-indexed ratio myth. 27:48 - Final verdict: How to stop fearing your Indexing Report. Resources Mentioned: Google Search Central: https://developers.google.com/search Google Search Console: https://search.google.com/search-console Search Central Live Events: https://developers.google.com/search/events Are you actively stressing over your non-indexed page counts, or are you tracking the big trend lines? Let us know in the comments! Episode transcript → https://goo.gle/sotr112-transcript Listen to more Search Off the Record → https://goo.gle/sotr-yt Subscribe to Google Search Channel → https://goo.gle/SearchCentral Search Off the Record is a podcast series that takes you behind the scenes of Google Search with the Search Relations team. #SOTRpodcast #SEO #GoogleSearch #SearchConsole Speakers: Martin Splitt, John Mueller

Sponsor detection runs nightly. Check back soon.
Based on semantic analysis of episode topics and host coverage, this show is a strong guest fit for executives in:
Industry fit is computed by PitchCentric using vector embeddings of the show's episode catalog.
Is this podcast yours and you'd like to remove or correct details? Request removal or email privacy@pitchcentric.com.
FAQs
If you have a concern about deliverability, AI quality, data privacy, or whether this will actually work for your specific situation, it's probably answered below.
Founder Solo gives you 50 AI pitches per month using the credit model (Standard pitches cost 1 credit, Enriched pitches cost 2). Founder Pro raises that to 200 credits per month and adds full Booking Probability access, unlimited Magic Match, Apollo enrichment credits, and data export capabilities. Both plans use the same credit system, so you can stretch your monthly budget further by using Standard-mode drafting.
Agency tiers have no base fee. You pay per managed client and per talent profile. Agency Standard is $199 per client per month; Agency Pro is $399 per client per month. Both add $39 per talent profile per month. Your own team's user seats are always free.
A talent profile represents one person (founder, executive, or spokesperson) you are booking onto podcasts. It includes their bio, topics, headshots, and outreach history. Team plans include 5 profiles; agency plans are pay-as-you-go.
Yes, at any time. Upgrades take effect immediately; downgrades apply at the end of the current billing period. Contact support if you need help migrating between plan families.
Every paid plan includes a 15-day free trial. Your card is saved at signup but you will not be charged until day 16. Cancel any time from your dashboard.
You keep access until the end of your current billing period. No charges after that. Your data is retained for 30 days in case you reactivate.
Yes. Select Annual on the pricing toggle and the discounted price is applied automatically at checkout. The annual price shown is the full year cost.
That is our Enterprise tier. Contact our sales team and we will build a custom plan with volume pricing, a dedicated account manager, and SLA guarantees.