Home / Other Search Topics / Search Theory / What Determines Which Website Pages Search Engines Crawl Regularly?

What Determines Which Website Pages Search Engines Crawl Regularly?

Aug 22, 2006 - 3:31 pm 0 — by Chris Boggs

Chris Boggs / Associate Editor

Chris has been a digital marketer since 2000, starting in-house in the insurance space and subsequently assuming growing leadership roles across large agencies and SEM firms. Chris has led SEO teams and developed Thought Leadership at two now-Publicis agencies in the 2000’s into early 2010’s, and since founding Web Traffic Advisors in 2014, consults business and agency clients helping to improve performance in SEO, Paid Search, and Paid Social Media. Chris supports organization teams and agencies to directly audit existing campaigns, lead the launch and management of new efforts, and analyze performance for actionable growth strategies. Chris also loves customized team training. Chris has current and recent experience serving major brands on various ecommerce platforms, B2B, Travel, Retail and other industries.

Since 2004 Chris has been a highly-rated speaker, moderator, and content advisor at conferences all over the world, covering search engine marketing, social media, analytics and evolving tactics. Chris is a regular speaker and moderator for the Pubcon conference series. He has served on the Global Board of Directors of SEMPO.org since 2006 and is a regular judge for the US Search Awards, given to high-performing brands, marketing agencies and consultants. Chris has published articles for various marketing publications through the years, although most his research and analysis these days is client-owned. Twitter: @boggles

248 Articles as chrisboggs

Filed Under Search Theory

Search engines use automated crawlers, also known as robots or spiders, to scour the Internet's content and add it to their indices. Once pages of a website are in an index that is used to provide search results, sites are revisited on a regular basis to determine if any new content has been added, or if there has been significant updates to currently indexed content. Since there is so much information available on the Internet, sites generally get re-crawled based on a variety of factors, including the frequency of content updates or even the command in a page's code that asks the spider to return every "X" days. However, except on occasion, spiders will not re-crawl the entire website.

A recent thread at Cre8asite Forums starts with a member asking "How do they do that?" He describes that he on occasion examines his log files to find varying degrees of robot activity, and asks how they determine how deep to dig. An initial answer by Moderator softplus offers some good ideas, and finishes with the thought that:

In the end, the main element I have seen for crawl frequency is page "value"; a page with good value is crawled more frequently than a page with little value... Even a static high-value page is crawled frequently, it doesn't make that much sense to me, but there must be reasoning behind it. Perhaps the frequency would be even higher if the content were to change frequently?

The member that asks the original question then poses the theory that the Google toolbar could be involved, with the crawl somehow directed towards pages with higher time spent by a visitor. The thread then diverges slightly into an interesting conversation about how Google Sitemaps works to help get pages crawled and indexed (now Google Webmaster Tools). Do you think you know why some pages are crawled and others not? Join the discussion at Cre8asite Forums.

Previous Story: A Proper Business Person's Use of AdSense?

Next Story: Brazil Fed Up With Google & Orkut

The content at the Search Engine Roundtable are the sole opinion of the authors and in no way reflect views of RustyBrick ®, Inc
Copyright © 1994-2025 RustyBrick ®, Inc. Web Development All Rights Reserved.
This work by Search Engine Roundtable is licensed under a Creative Commons Attribution 3.0 United States License. Creative Commons License and YouTube videos under YouTube's ToS.

What Determines Which Website Pages Search Engines Crawl Regularly?

Chris Boggs / Associate Editor

Popular Categories

The Pulse of the search community

Search Video Recaps

Most Recent Articles

Daily Search Forum Recap: April 18, 2025

Search News Buzz Video Recap: Google Ruled A Monopoly Again, Heated Volatility, Google ccTLD Change, Ads Safety Report & AI Overviews

Google Made Search Much Faster - But How Much Faster?

New Google Merchant Center Popular Products Report

Google: No Positive SEO Effect From .esports Domain

Google Ads PMax Age Exclusions Rolling Out