Google: Do Not Block GoogleBot From Crawling 404s

Jul 15, 2020 - 7:57 am 0 by

Google Blocked

John Mueller of Google said it would be "a really bad idea which will cause all sorts of problems" if you block Google or other search engines from crawling pages that return a 404 server status code. He said "billions of 404 pages are crawled every day" by Google and it is normal.

One webmaster wrote that his "website automatically blocks user agents that get more than 10 404 errors, including Googlebot, so that's a problem." John responded to that that this is a really bad idea, he said "That sounds like a really bad idea which will cause all sorts of problems.. You can't avoid that Googlebot & all other search engines will run into 404s. Crawling always includes URLs that were previously seen to be 404."

He said in a different tweet, the same day, "Billions of 404 pages are crawled every day - it's a normal part of the web, it's the proper way to signal that a URL doesn't exist. That's not something you need to, or can, suppress."

So while you can fix your 404 pages through other means, automatically blocking Google from accessing 404 pages without knowing how Google is accessing those pages can be a really bad idea.

Forum discussion at Twitter.

 

Popular Categories

The Pulse of the search community

Follow

Search Video Recaps

 
- YouTube
Video Details More Videos Subscribe to Videos

Most Recent Articles

Google Updates

Google November 2024 Core Update Heats Up This Weekend

Nov 17, 2024 - 8:10 am
Search Forum Recap

Daily Search Forum Recap: November 15, 2024

Nov 15, 2024 - 10:00 am
Search Video Recaps

Search News Buzz Video Recap: Google November 2024 Core Update, AI Overview Hyperlinks, SEO, Ads, AdSense & More

Nov 15, 2024 - 8:01 am
Google Maps

Google Maps Search For Products Nearby Carousel

Nov 15, 2024 - 7:51 am
Google

iPhone Gets Native Google Gemini App

Nov 15, 2024 - 7:41 am
Google

Google Chrome To Spotlight Merchant Center Promotions

Nov 15, 2024 - 7:31 am
Previous Story: Google Discover AMP Articles Going To Main Canonical URL?