Does “authority” mean quality? predicting expert quality ratings of Web documents
Published 1 July 2000
Brian Amento, Loren Terveen, Will Hill
Citations227
Generate an AI Snapshot to get a quick, structured summary of this paper.
Study Snapshot
ObjectiveStudy objective
MethodsResearch methodology
PopulationPopulation studied
Sample sizeSample sizes
OutcomesStudy outcomes here
ResultsStudy results comes here
LimitationsResearch study limitations comes here
A concise AI-generated summary of the paper will appear here once you click Generate AI Snapshot.
TL;DR
An experimental evaluation of link analysis algorithms for their potential to identify high quality items using a dataset of web documents rated for quality by human topic experts found link-based metrics did a good job of picking out high-quality items.
Abstract
For many topics, the World Wide Web contains hundreds or thousands of relevant documents of widely varying quality. Users face a daunting challenge in identifying a small subset of documents worthy of their attention.
Keywords
Computer SciencePhysics and Astronomy
The PageRank Citation Ranking : Bringing Order to the Web
12,645 Citations1999Lawrence M. Page, Sergey Brin +2 more
This paper describes PageRank, a mathod for rating Web pages objectively and mechanically, effectively measuring the human interest and attention devoted to them, and shows how to efficiently compute PageRank for large numbers of pages.
Journal of the ACMAuthoritative sources in a hyperlinked environment
9,060 Citations1999Jon Kleinberg
This work proposes and test an algorithmic formulation of the notion of authority, based on the relationship between a set of relevant authoritative pages and the set of “hub pages” that join them together in the link structure, and has connections to the eigenvectors of certain matrices associated with the link graph.
Computer Networks and ISDN SystemsAutomatic resource compilation by analyzing hyperlink structure and associated text
700 Citations1998Soumen Chakrabarti, Byron Dom +4 more
An evaluation of ARC suggests that the resources found by ARC frequently fare almost as well as, and sometimes better than, lists of resources that are manually compiled or classified into a topic.
ACM SIGIR ForumImproved Algorithms for Topic Distillation in a Hyperlinked Environment
677 Citations2017Krishna Bharat, Monika Henzinger
This paper addresses the problem of topic distillation on the World Wide Web, namely, given a typical user query to find quality documents related to the query topic, by augmenting a previous connectivity analysis based algorithm with content analysis.
Improved algorithms for topic distillation in a hyperlinked environment
439 Citations1998Krishna Bharat, Monika Henzinger
This paper addresses the problem of topic distillation on the World Wide Web, namely, given a typical user query to find quality documents related to the query topic, by augmenting a previous connectivity analysis based algorithm with content analysis.
Silk from a sow's ear
382 Citations1996Peter Pirolli, James E. Pitkow +1 more
This paper presents the exploration into techniques that utilize both the topology and textual similarity between items as well as usage data collected by servers and page meta-information lke title and size.
The WebBook and the Web Forager
342 Citations1996Stuart K. Card, George G. Robertson +1 more
This paper presents two related designs with which to evolve the Web and its clients, a 3D interactive book of HTML pages and the Web Forager, an application that embeds the WebBook and other objects in a hierarchical 3D workspace.
Scatter/gather browsing communicates the topic structure of a very large text collection
196 Citations1996Peter Pirolli, Patricia Schänk +2 more
The results suggest that Scatter/Gather induces a more coherent conceptual image of a text collection, a richer vocabulary for constructing search queries, and communicates the distribution of relevant documents over clusters of documents in the collection.
SenseMaker
167 Citations1997Michelle Baldonado, Terry Winograd
The design and implementation of SenseMaker is described, an interface for information exploration across heterogeneous sources, and how SenseMaker supports the context-driven evolution of a user's interests is discussed.
Life, death, and lawfulness on the electronic frontier
152 Citations1997James E. Pitkow, Peter Pirolli
To facilitate users’ ability to make sense of large collections of hypertext, two new techniques for inducing clusters of related documents on the World Wide Web are presented.
ACM Transactions on Computer-Human InteractionConstructing, organizing, and visualizing collections of topically related Web resources
101 Citations1999Loren Terveen, Will Hill +1 more
The auditorium visualization, augmented with drill-down capabilities to explore site profile data, helps users to find high-quality sites as well as sites that serve a particular function.
An empirical evaluation of user interfaces for topic management of Web sites
38 Citations1999Brian Amento, Will Hill +3 more
TopicShop includes a webcrawler that discovers relevant web sites and builds site profiles, and user interfaces for exploring and organizing sites, and provides the key to these results, as users exploited it to identify themost promising sites quickly and easily.
VTechWorks (Virginia Tech)User interfaces for topic management of web sites
2 Citations2001Brian Amento, Deborah Hix
The site profile data that TopicShop provide—in particular, the number of pages on a site and the number of other sites that link to it—were the key to these results, as users exploited them to identify the most promising sites quickly and easily.
