Conceptualizing documentation on the Web: An evaluation of different heuristic‐based models for counting links between university Web sites

Mike Thelwall

doi:10.1002/asi.10135

Loading next page...

References (48)

L. Egghe (2000)
New informetric aspects of the Internet: some reflections - many problems
Journal of Information Science, 26
Georg Rehm (2002)
Towards automatic Web genre identification: a corpus-based approach in the domain of academia by example of the Academic's Personal Homepage
Proceedings of the 35th Annual Hawaii International Conference on System Sciences
J. Kleinberg (1999)
Authoritative sources in a hyperlinked environment
S. Harter, C. Ford (2000)
Web-based analyses of E-journal impact: Approaches, problems, and issues
J. Am. Soc. Inf. Sci., 51
A. Broder, Ravi Kumar, F. Maghoul, P. Raghavan, S. Rajagopalan, Raymie Stata, A. Tomkins, J. Wiener (2000)
Graph structure in the Web
Comput. Networks, 33
L. Leydesdorff, Michaela Curran (2000)
Mapping university-industry-government relations on the Internet: The construction of indicators for a knowledge-based economy.
, 4
Neil Jacobs (2001)
Information technology and interests in scholarly communication: A discourse analysis
J. Assoc. Inf. Sci. Technol., 52
Lennart Björneborn (2001)
Small-world linkage and co-linkage
Proceedings of the 12th ACM conference on Hypertext and Hypermedia
L. Cui (1999)
Rating Health Web sites using the principles of Citation Analysis: A Bibliometric Approach
Journal of Medical Internet Research, 1
S. Harnad, L. Carr (2000)
Integrating, Navigating and Analyzing Eprint Archives Through Open Citation Linking (the OpCit Project)
R. Larson (1996)
Bibliometrics of the World Wide Web: An Exploratory Analysis of the Intellectual Structure of Cyberspace
, 33
E. Davenport, B. Cronin (2000)
The citation network as a prototype for representing trust.
M. Thelwall (2001)
A web crawler design for data mining
Journal of Information Science, 27
P. Ingwersen (1998)
The calculation of web impact factors
J. Documentation, 54
M. Thelwall (2002)
Methodologies for crawler based Web surveys
Internet Res., 12
A. Goodrum, K. McCain, S. Lawrence, C. Giles (2001)
Scholarly publishing in the Internet age: a citation analysis of computer science literature
Inf. Process. Manag., 37
M. Thelwall (2000)
Web impact factors and search engine coverage
J. Documentation, 56
E. Garfield, B. Cronin, Helen Atkins (2000)
The web of knowledge: a festschrift in honor of Eugene Garfield
(1996)
Studying research collaboration using coauthorships, Scientometrics
R. Rousseau (1998)
Daily time series of common single word searches in AltaVista and NorthernLight
R. Rousseau (1997)
Sitations: an exploratory study
M. Bates, Shaojun Lu (1997)
AN EXPLORATORY PROFILE OF PERSONAL HOME PAGES: CONTENT, DESIGN, METAPHORS
, 21
M. Thelwall (2001)
A publicly accessible database of UK university website links and a discussion of the need for human
Rong Tang, M. Thelwall (2005)
Exploring the pattern of links between Chinese university Web sites
Georg Rehm (2002)
Towards Automatic Web Genre Identification
Blaise Cronin (2001)
Bibliometrics and beyond: some thoughts on web-based citation analysis
Journal of Information Science, 27
M. Thelwall (2002)
Evidence for the existence of geographic trends in university Web site interlinking
J. Documentation, 58
A. Dillon, Misha Vaughan (1997)
â Itâ s the journey and the destinationâ : Shape and the emergent property of genre in evaluating digital documents
(1994)
The impact factor, Current Contents, June 20
M. Thelwall (2002)
A comparison of sources of links for academic Web impact factor calculations
J. Documentation, 58
H. Chu, Shaoyi He, M. Thelwall (2002)
Library and Information Science Schools in Canada and USA: A Webometric Perspective
Journal of Education for Library and Information Science, 43
Lennart Björneborn, P. Ingwersen (2004)
Perspective of webometrics
Scientometrics, 50
(2000)
Mesure de l'impact des sites Web : le Web Impact Factor. L'exemple des CHU français. Revue du Praticien -Médecine Générale
Wouter Mettrop, P. Nieuwenhuysen (2001)
Internet search engines - fluctuations in document accessibility
J. Documentation, 57
B. Cronin, H. Snyder, H. Rosenbaum, Anna Martinson, E. Callahan (1998)
Invoked on the Web
J. Am. Soc. Inf. Sci., 49
Alastair Smith, M. Thelwall (2002)
Web Impact Factors for Australasian universities
Scientometrics, 54
John O'Leary, A. Hindmarsh, B. Kingston (2006)
The Times Good University Guide 2007
Stephanie Haas, Erika Grams (1999)
Structure : A Discussion of Four Questions Arising from a Content Analysis of Web Pages
M. Thelwall (2002)
A research and institutional size-based model for national university Web site interlinking
J. Documentation, 58
M. Thelwall (2001)
Extracting macroscopic information from Web links
J. Assoc. Inf. Sci. Technol., 52
M. Thelwall (2001)
Results from a web impact factor crawler
J. Documentation, 57
(2001)
Necessary data filtering and editing in webometric link structure analysis
S. Brin, Lawrence Page (1998)
The Anatomy of a Large-Scale Hypertextual Web Search Engine
Comput. Networks, 30
Kevin Crowston, M. Williams (1997)
Reproduced and emergent genres of communication on the World-Wide Web
Proceedings of the Thirtieth Hawaii International Conference on System Sciences, 6
(2002)
Profile Available: http://dandini.cranfield.ac.uk/profile
Alastair Smith (1999)
A Tale of Two Web Spaces: Comparing Sites Using Web Impact Factors.
Journal of Documentation, 55
R. Kling, G. McKim (1999)
Not just a matter of time: Field differences and the shaping of electronic media in supporting scientific communication
J. Am. Soc. Inf. Sci., 51
M. Wikgren (2001)
Health discussions on the Internet
Library & Information Science Research, 23

Publisher: Wiley
Copyright: Copyright © 2002 Wiley Subscription Services, Inc., A Wiley Company
ISSN: 2330-1635
eISSN: 2330-1643
DOI: 10.1002/asi.10135
Publisher site: See Article on Publisher Site

Abstract

All known previous Web link studies have used the Web page as the primary indivisible source document for counting purposes. Arguments are presented to explain why this is not necessarily optimal and why other alternatives have the potential to produce better results. This is despite the fact that individual Web files are often the only choice if search engines are used for raw data and are the easiest basic Web unit to identify. The central issue is of defining the Web “document”: that which should comprise the single indissoluble unit of coherent material. Three alternative heuristics are defined for the educational arena based upon the directory, the domain and the whole university site. These are then compared by implementing them on a set of 108 UK university institutional Web sites under the assumption that a more effective heuristic will tend to produce results that correlate more highly with institutional research productivity. It was discovered that the domain and directory models were able to successfully reduce the impact of anomalous linking behavior between pairs of Web sites, with the latter being the method of choice. Reasons are then given as to why a document model on its own cannot eliminate all anomalies in Web linking behavior. Finally, the results from all models give a clear confirmation of the very strong association between the research productivity of a UK university and the number of incoming links from its peers' Web sites.

Journal

Journal of the Association for Information Science and Technology – Wiley

Published: Oct 1, 2002

Get 20M+ Full-Text Papers For Less Than $1.50/day. Start a 14-Day Trial for You or Your Team.

Learn More →

Conceptualizing documentation on the Web: An evaluation of different heuristic‐based models for counting links between university Web sites

Conceptualizing documentation on the Web: An evaluation of different heuristic‐based models for counting links between university Web sites

Get 20M+ Full-Text Papers For Less Than $1.50/day. Start a 14-Day Trial for You or Your Team.

Learn More →

Conceptualizing documentation on the Web: An evaluation of different heuristic‐based models for counting links between university Web sites

Conceptualizing documentation on the Web: An evaluation of different heuristic‐based models for counting links between university Web sites

References (48)

Abstract

Journal

Recommended Articles

There are no references for this article.

Our policy towards the use of cookies