Featured Posts

A Chain Mail Victim’s Hilarious MailA Chain Mail Victim’s Hilarious Mail Hi Pals, I am writing this post after a very long time because I was suffering from my health problems and was not able to sit in front of my computer. But I am sure that this post will give you...

Readmore

Indian Rupee Got its Own SymbolIndian Rupee Got its Own Symbol This post is not related to seo or Blogging, but I thought I shall share it with you all, as I am feeling very proud that Indian Currency has got its own symbol. The symbol combines the Devnagri “Ra”...

Readmore

How to Block an IP from Accessing Your WebsiteHow to Block an IP from Accessing Your Website Some days before i was checking logs of one of my website related to travel and i found that an IP was sucking a lot of bandwidth from my hosting account. But the website (the IP) was not sending any visitor...

Readmore

Namecheap Coupon Code for July 2010Namecheap Coupon Code for July 2010 The NameCheap.Com (aff link) is my favorite domain reseller registrar and I book all my domains from them, Infact I have registered over 70 domains with them, because in my opinion they have the cheapest...

Readmore

Adsense Revenue Share Secret is Finally OutAdsense Revenue Share Secret is Finally Out The great secret of Adsense revenue share with Google and publishers is finally out and the company took this step for greater transparency in their operations. The percentage of revenue which Google...

Readmore

  • Prev
  • Next

Google Basics: How Google Works

Posted on : 30-09-2009 | By : rituraj | In : Online Business, Search Engine Optimization

0

I got a very informative article about how Google crawl and index pages andHow google works how it serves results for a user query. In this article you can get insight of the basics of the mechanism of search engine and can find hints to optimize your WebPages better. Hope you all will find it useful.

When you sit down at your computer and do a Google search, you’re almost instantly presented with a list of results from all over the web. How does Google find web pages matching your query, and determine the order of search results?

In the simplest terms, you could think of searching the web as looking in a very large book with an impressive index telling you exactly where everything is located. When you perform a Google search, our programs check our index to determine the most relevant search results to be returned (”served”) to you.

The three key processes in delivering search results to you are:

Crawling

Crawling is the process by which Googlebot discovers new and updated pages to be added to the Google index.

We use a huge set of computers to fetch (or “crawl”) billions of pages on the web. The program that does the fetching is called Googlebot (also known as a robot, bot, or spider). Googlebot uses an algorithmic process: computer programs determine which sites to crawl, how often, and how many pages to fetch from each site.

Google’s crawl process begins with a list of web page URLs, generated from previous crawl processes, and augmented with Sitemap data provided by webmasters. As Googlebot visits each of these websites it detects links on each page and adds them to its list of pages to crawl. New sites, changes to existing sites, and dead links are noted and used to update the Google index.

Google doesn’t accept payment to crawl a site more frequently, and we keep the search side of our business separate from our revenue-generating AdWords service.

Indexing

Googlebot processes each of the pages it crawls in order to compile a massive index of all the words it sees and their location on each page. In addition, we process information included in key content tags and attributes, such as Title tags and ALT attributes. Googlebot can process many, but not all, content types. For example, we cannot process the content of some rich media files or dynamic pages.

Serving results

When a user enters a query, our machines search the index for matching pages and return the results we believe are the most relevant to the user. Relevancy is determined by over 200 factors, one of which is the PageRank for a given page. PageRank is the measure of the importance of a page based on the incoming links from other pages. In simple terms, each link to a page on your site from another site adds to your site’s PageRank. Not all links are equal: Google works hard to improve the user experience by identifying spam links and other practices that negatively impact search results. The best types of links are those that are given based on the quality of your content.

In order for your site to rank well in search results pages, it’s important to make sure that Google can crawl and index your site correctly. Our Webmaster Guidelines outline some best practices that can help you avoid common pitfalls and improve your site’s ranking.

Google’s Related Searches, Spelling Suggestions, and Google Suggest features are designed to help users save time by displaying related terms, common misspellings, and popular queries. Like our google.com search results, the keywords used by these features are automatically generated by our web crawlers and search algorithms. We display these suggestions only when we think they might save the user time. If a site ranks well for a keyword, it’s because we’ve algorithmically determined that its content is more relevant to the user’s query.

You Can Visit Source Here at Google Webmaster Support

Bookmark & Share!!!
[Connotea] [del.icio.us] [Digg] [diigo] [Facebook] [Fark] [Faves] [LinkedIn] [Netvouz] [Reddit] [Squidoo] [StumbleUpon] [Technorati] [Twitter] [Yahoo!]

Popularity: 13% [?]

Write a comment