Persistence of Web references in scientific research
ComputerPublished 1 March 2001
Sandra Lawrence, D.M. Pennock, Gary William Flake, Robert Krovetz, Frans Coetzee, Eric Glover
Citations233
SJR quartileQ2
SJR score0.55
SNIP1.02
Generate an AI Snapshot to get a quick, structured summary of this paper.
Study Snapshot
ObjectiveStudy objective
MethodsResearch methodology
PopulationPopulation studied
Sample sizeSample sizes
OutcomesStudy outcomes here
ResultsStudy results comes here
LimitationsResearch study limitations comes here
A concise AI-generated summary of the paper will appear here once you click Generate AI Snapshot.
TL;DR
It is argued that although few critical resources have been lost to date, new strategies to manage Internet resources and improved citation practices are necessary to minimize the future loss of information.
Abstract
The lack of persistence of Web references has called into question the increasingly common practice of citing URLs in scientific papers. It is argued that although few critical resources have been lost to date, new strategies to manage Internet resources and improved citation practices are necessary to minimize the future loss of information.
Keywords
Computer Science
NatureAccessibility of information on the web
1,356 Citations1999Steve Lawrence, C. Lee Giles
As the web becomes a major communications medium, the data on it must be made more accessible, and search engines need to make the data more accessible.
ComputerDigital libraries and autonomous citation indexing
614 Citations1999Sandra Lawrence, C. Lee Giles +1 more
Digital libraries incorporating ACI can help organize scientific literature and may significantly improve the efficiency of dissemination and feedback and speed the transition to scholarly electronic publishing.
IEEE Internet ComputingContext and page analysis for improved Web search
184 Citations1998Sandra Lawrence, C. Lee Giles
The paper discusses the features of the NECI metasearch engine and suggests ways to improve the efficiency of Web searches by downloading and analyzing each document and then displaying results that show the query terms in concert.
Functional Requirements for Uniform Resource Names
182 Citations1994Karen Sollins, Larry Masinter
This document specifies a minimum set of requirements for a kind of Internet resource identifier known as Uniform Resource Names (URNs) and provides information for the Internet community.
Indexing and retrieval of scientific literature
112 Citations1999Steve Lawrence, Kurt Bollacker +1 more
This paper discusses the creation of digital libraries of scientific literature on the web, including the efficient location of articles, full-text indexing of the articles, autonomous citation indexing, information extraction, display of query-sensitive summaries and citation context, hubs and authorities computation.
Towards an archival Intermemory
108 Citations2002Andrew V. Goldberg, P.N. Yianilos
This paper presents a framework for the design of an Intermemory, and considers certain aspects of the design in greater detail, in particular the aspects of addressing, space efficiency, and redundant coding are discussed.
The Hyper-G Network Information System
62 Citations1996Keith Andrews, Frank Kappe +1 more
Computer Networks and ISDN SystemsFixing the “broken-link” problem: The W3Objects approach
51 Citations1996David B. Ingham, S.J. Caughey +1 more
A model for the provision of referential integrity for Web resources which supports resource migration and tolerates site and communication failures is presented, which is object-oriented, highly flexible, completely distributed, and does not require any global administration.
Robust Hyperlinks Cost Just Five Words Each
50 Citations2000Thomas A. Phelps, Robert Wilensky
This paper proposes robust hyperlinks as a solution to the problem of broken hyperlinks, a URL augmented with a small "signature", computed from the referenced document, which can be submitted as a query to web search engines to locate the document.
USENIX Annual Technical ConferencePermanent web publishing
48 Citations2000David S. H. Rosenthal, Vicky Reich
LOCKSS (Lots Of Copies Keep Stuff Safe) is a prototype of a system to preserve access to scientific journals published on the Web that, unlike normal systems, has far more replicas than would be required just to survive the anticipated failures.
Intelligent Agents for Web-based Tasks: An Advice-Taking Approach
25 Citations1998Jude Shavlik, Tina Eliassi‐Rad
The architecture provides an appealing middle ground between nonadaptive agent programming languages and systems that solely learn user preferences from the user’s ratings of pages, and how advice is mapped into neural network implementations of the two functions.
Computer Networks and ISDN SystemsAuthor-oriented link management
18 Citations1996Michael Creech
This paper describes a link management technique that helps authors ensure the consistency of their content, called the change log table/web-walk (CLT/WW) approach, which is based on web operations that log the gross changes made by authors to their web pages.
A formal approach to analyzing the browsing semantics of hypertext
7 Citations1994P.M.E. De Bra, G.J.P.M. Houben +1 more
This paper combines the operations offered by the hypertext system (or presentation layer) with the link structure of the document, in order to characterize the browsing semantics of document and system as a unity.
