Get 20M+ Full-Text Papers For Less Than $1.50/day. Start a 14-Day Trial for You or Your Team.

Learn More →

Conceptualizing documentation on the Web: An evaluation of different heuristic‐based models for counting links between university Web sites

Conceptualizing documentation on the Web: An evaluation of different heuristic‐based models for... All known previous Web link studies have used the Web page as the primary indivisible source document for counting purposes. Arguments are presented to explain why this is not necessarily optimal and why other alternatives have the potential to produce better results. This is despite the fact that individual Web files are often the only choice if search engines are used for raw data and are the easiest basic Web unit to identify. The central issue is of defining the Web “document”: that which should comprise the single indissoluble unit of coherent material. Three alternative heuristics are defined for the educational arena based upon the directory, the domain and the whole university site. These are then compared by implementing them on a set of 108 UK university institutional Web sites under the assumption that a more effective heuristic will tend to produce results that correlate more highly with institutional research productivity. It was discovered that the domain and directory models were able to successfully reduce the impact of anomalous linking behavior between pairs of Web sites, with the latter being the method of choice. Reasons are then given as to why a document model on its own cannot eliminate all anomalies in Web linking behavior. Finally, the results from all models give a clear confirmation of the very strong association between the research productivity of a UK university and the number of incoming links from its peers' Web sites. http://www.deepdyve.com/assets/images/DeepDyve-Logo-lg.png Journal of the Association for Information Science and Technology Wiley

Conceptualizing documentation on the Web: An evaluation of different heuristic‐based models for counting links between university Web sites

Loading next page...
 
/lp/wiley/conceptualizing-documentation-on-the-web-an-evaluation-of-different-ydKTdhwTu2

References (48)

Publisher
Wiley
Copyright
Copyright © 2002 Wiley Subscription Services, Inc., A Wiley Company
ISSN
2330-1635
eISSN
2330-1643
DOI
10.1002/asi.10135
Publisher site
See Article on Publisher Site

Abstract

All known previous Web link studies have used the Web page as the primary indivisible source document for counting purposes. Arguments are presented to explain why this is not necessarily optimal and why other alternatives have the potential to produce better results. This is despite the fact that individual Web files are often the only choice if search engines are used for raw data and are the easiest basic Web unit to identify. The central issue is of defining the Web “document”: that which should comprise the single indissoluble unit of coherent material. Three alternative heuristics are defined for the educational arena based upon the directory, the domain and the whole university site. These are then compared by implementing them on a set of 108 UK university institutional Web sites under the assumption that a more effective heuristic will tend to produce results that correlate more highly with institutional research productivity. It was discovered that the domain and directory models were able to successfully reduce the impact of anomalous linking behavior between pairs of Web sites, with the latter being the method of choice. Reasons are then given as to why a document model on its own cannot eliminate all anomalies in Web linking behavior. Finally, the results from all models give a clear confirmation of the very strong association between the research productivity of a UK university and the number of incoming links from its peers' Web sites.

Journal

Journal of the Association for Information Science and TechnologyWiley

Published: Oct 1, 2002

There are no references for this article.