smallbusiness.itworld.com
  Search  
Small Business Home Page Small Business Webcasts Small Business White Papers Small Business Newsletters Small Business News Small Business Topic Map Careers ITworld Voices ITwhirled The site for Small and Medium Business from ITworld.com

Inside Google Sitemaps

ITworld.com 2/1/06

Google Sitemaps is a program that lets any site developer publish a map of their site and submit it to Google for indexing. It's designed to let companies guide Google's crawlers and should help get pages indexed more quickly and thoroughly.

The program is free, and given the importance of Google traffic to most sites, it's worth taking a look at how it works and what tools are available to support it.

Why Sitemaps Matter

It can be frustrating to see how long it takes for web pages to get indexed at sites like Google. If your pages don't show up in Google, you're probably losing out on traffic and sales or advertising revenue.

Sitemaps are a way to help Google index your site. While Google says that this won't raise your page ranking, it may mean that your site gets more traffic, simply because your pages are more thoroughly indexed.

Google has published a case study on Sitemaps looking at the experience of Interactive Sites, a developer of web-based products for the hospitality industry. The company found Sitemaps easy to implement, and effective at increasing traffic to their sites.

According to John Blayter, Interactive Sites' Director of Engineering, "We were incredibly impressed with how easy Sitemaps was to integrate into our CMS." Blayter and his team integrated Sitemaps into the company's CMS in a single day and placed it on 60 client websites.

Interactive Sites' implementation of Sitemaps improved both the coverage and freshness of content in Google's index. Since implementing Sitemaps, the company's clients have benefited from an average increase of 125 percent in the number of indexed pages - and in some cases more than 240 percent.

For many companies, an increase of 125 percent in the number of pages indexed could result in a big jump in site traffic.

What are Sitemaps?

A Sitemap is an XML file that resides on your website that is updated regularly with any changes and additions to your site.

The file is in a fairly simple format. In Sitemap XML files, the main element, urlset, encloses a collection of URLs. For each page on your site that you want to include in the Sitemap file, there is a url element with one required child element, the loc, which is just the page's URL. So, in its simplest form, a Sitemap is just a big list of the URLs at your site.

There are also several optional child elements for the url element, lastmod, changefreq, and priority:
* lastmod indicates when the URL was last update
* changefreq indicates how often the page is typically updated
* priority indicates the importance of this URL relative to the other URLs at your site.

Tools for Supporting Sitemap

There are many tools available to support using Sitemaps. First on the list would be tools for validating the Sitemap file. Because Sitemaps are XML files, you can use standard XML tools to validate them. Google has published XML schema defining the elements and attributes that can be used in Sitemaps (see Resources).

Google provides a script that is designed to help site owners create Sitemaps, Sitemap Generator. It's a free python script, downloadable via Sourceforge (see Resources). Many third-party tools have also emerged to support the creation of Sitemaps, including standalone scripts and updates to content management systems.

Google also provides a Web-based tool for checking on the status of your sites within their system. It lets you see if Google has read your Sitemap, and if it was successful or if there was an error.

You can also Verify your site, which lets you look at more detailed statistics. Google provides you with a unique file name for your site. It looks something like this: google8ad97837ed3875e83.html. You create an empty file with the name that Google provides and put it at the root level of your Web site. Then you log into Google and click the Verify button within your account. This process lets Google know that you control the site.

Once you have verified your site, Google provides additional information:
* Sitemap details and errors
* Indexing information about your site
* Query stats about your site
* Crawl stats about your site
* Page analysis of your site
* URLs from your site we were unable to crawl, and why we couldn't crawl them

Google Sitemaps provides a way for sites to ensure that they are thoroughly indexed. This, and the Web-based system Google provides for understanding the indexing of your site, make Sitemaps an important tool for web developers and site owners.

ADDITIONAL RESOURCES

Sitemap Protocol

Sitemap Schema
For Sitemaps
For Sitemap index files

Getting started with Google sitemaps

Sitemap Generator

Google Site Overview (Google account required)

On this topic

 




Sponsored Links

Used and Refurbished Cisco Routers
Purchase Your Routers From Network Liquidators. Savings of Up to 90% with a Lifetime Warranty!
SOLVE SUPPORT ISSUES on the First Call!
REMOTELY CONTROL AND CONFIGURE SYSTEMS. Easily install applications, updates. All from your Desktop!
TAKE CONTROL OF REMOTE COMPUTERS
Support, configure and install applications and updates remotely for greater efficiency.
RESOLVE SUPPORT ISSUES from your Desktop!
Minimize downtime with a remote support solution that lets you resolve issues right from the desktop
See how EASY REMOTE SUPPORT can be. Try WebEx FREE!
DELIVER SUPPORT MORE EFFICIENTLY. Remotely Control Applications. Leap Securely through Firewalls!
» Buy a link now

Advertisements
Sponsored links
Bring harmony to your mix of UNIX-Linux-Windows computing environments
Top 5 Reasons to Combine App Performance and Security
KODAK i1400 Series Scanners stand up to the challenge
Locate Hidden Software on business PCs with this free tool
 Home   Software and services  Web site management
www.itworld.com    open.itworld.com     security.itworld.com     smallbusiness.itworld.com
storage.itworld.com     utilitycomputing.itworld.com     wireless.itworld.com

 
Contact Us   About Us   Privacy Policy    Terms of Service   Reprints  

CIO   Computerworld   CSO   GamePro   Games.net   IDG Connect   IDG World Expo   Infoworld   ITworld   JavaWorld   LinuxWorld  MacUser   Macworld   Network World   PC World   Playlist  

Copyright © Computerworld, Inc. All rights reserved

Reproduction in whole or in part in any form or medium without express written permission of Computerworld Inc. is prohibited. Computerworld and Computerworld.com and the respective logos are trademarks of International Data Group Inc.