Home / Google News / Google SEO / Peeking Into Google Reveals More of Google's Architecture

Peeking Into Google Reveals More of Google's Architecture

Mar 13, 2006 - 7:55 am 2 — by Barry Schwartz

Filed Under Google Search Engine Optimization

An article at InternetNews.com from March 2nd, has Google vice president of operations and vice president of engineering, Urs Hoelzle revealing some of the "behind-the-scenes tour of Google's architecture."

Bill Slawski at Cre8asite Forums created a thread on this article named Google's architecture, Informative news story where he pulled out a couple quotes.

Google replicates the Web pages it caches by splitting them up into pieces it calls "shards." The shards are small enough that several can fit on one machine. And they're replicated on several machines, so that if one breaks, another can serve up the information. The master index is also split up among several servers, and that set also is replicated several times. The engineers call these "chunk servers."

The company also is applying machine learning to its system to give better results. Theoretically, he said, if someone searches for "Bay Area cooking class," the system should know that "Berkeley courses: vegetarian cuisine" is a good match even though it contains none of the query words.
To do this, the system tries to cluster concepts into "reasonably coherent" subclusters that seem related. These clusters, some tiny and some huge, are named automatically. Then, when a query comes in, the system produces a probability score for the various clusters. This kind of machine learning has had little success in academic trials, Hoelzle said, because they didn't have enough data. "If you have enough data, you get reasonably good answers out of it."

The article is definitely worth a read and then join the forum discussion at Cre8asite Forums.

Previous Story: Possible Yahoo! Search Update

Next Story: Rumors of Google Hiring 15 Year Old Not True

The content at the Search Engine Roundtable are the sole opinion of the authors and in no way reflect views of RustyBrick ®, Inc
Copyright © 1994-2025 RustyBrick ®, Inc. Web Development All Rights Reserved.
This work by Search Engine Roundtable is licensed under a Creative Commons Attribution 3.0 United States License. Creative Commons License and YouTube videos under YouTube's ToS.

Peeking Into Google Reveals More of Google's Architecture

Barry Schwartz / Executive Editor

Popular Categories

The Pulse of the search community

Search Video Recaps

Most Recent Articles

Daily Search Forum Recap: April 18, 2025

Search News Buzz Video Recap: Google Ruled A Monopoly Again, Heated Volatility, Google ccTLD Change, Ads Safety Report & AI Overviews

Google Made Search Much Faster - But How Much Faster?

New Google Merchant Center Popular Products Report

Google: No Positive SEO Effect From .esports Domain

Google Ads PMax Age Exclusions Rolling Out