Tapping the power of text mining
Communications of the ACMPublished 1 September 2006
Weiguo Fan, Linda Wallace, Stephanie Rich, Zhongju Zhang
Citations433
SJR quartileQ1
SJR score1.15
SNIP3.34
Generate an AI Snapshot to get a quick, structured summary of this paper.
Study Snapshot
ObjectiveStudy objective
MethodsResearch methodology
PopulationPopulation studied
Sample sizeSample sizes
OutcomesStudy outcomes here
ResultsStudy results comes here
LimitationsResearch study limitations comes here
A concise AI-generated summary of the paper will appear here once you click Generate AI Snapshot.
TL;DR
Sifting through vast collections of unstructured or semistructured data beyond the reach of data mining tools, text mining tracks information sources, links isolated concepts in distant documents, maps relationships between activities, and helps answer questions.
Abstract
Sifting through vast collections of unstructured or semistructured data beyond the reach of data mining tools, text mining tracks information sources, links isolated concepts in distant documents, maps relationships between activities, and helps answer questions.
Keywords
Computer ScienceBiochemistry, Genetics and Molecular Biology
A Comparative Study on Feature Selection in Text Categorization
4,766 Citations1997Yiming Yang, Jan Pedersen
DF thresholding, the simplest method with the lowest cost in computation, can be reliably used instead of IG or CHI when the computation of these measures are too expensive, and strong correlations between the DF, IG and CHI values of a term are found.
Journal of the American Society for Information ScienceTwo medical literatures that are logically but not bibliographically connected
197 Citations1987Don R. Swanson
It is demonstrated that certain unintended logical connections within the scientific literature are unmarked by reference citations or other bibliographic clues, and may be inherently and peculiarly difficult to solve because there are virtually no references in either literature to the other.
A system for automatic personalized tracking of scientific literature on the Web
121 Citations1999Kurt Bollacker, Steve Lawrence +1 more
A system as part of the CiteSeer digital library project for automatic tracking of scientific literature that is relevant to a user’s research interests that is able to track and recommend topically relevant papers even when keyword based query profiles fail.
Communications of the ACMEmerging scientific applications in data mining
79 Citations2002Jiawei Han, Russ B. Altman +3 more
Automated, scalable systems would reveal and help exploit the deeper meanings in scientific data, especially in biomedical engineering, telecommunications, geospatial exploration, and climate and Earth ecosystem modeling.
ACM Transactions on Internet TechnologyLiterature-based discovery on the World Wide Web
59 Citations2002Michael Gordon, Robert Lindsay +1 more
This work suggests that literature-based discovery can be fruitful in areas other than medicine, and that it can support individuals seeking to draw together ideas from various areas of inquiry, even if such connections have been previously made by others.
Journal of the American Society for Information Science and TechnologyGetting answers to natural language questions on the Web
34 Citations2002Dragomir Radev, Kelsey Libner +1 more
This work identified the best-performing search engines overall for factual natural language questions and found performance differences depending on the domain of factual question asked.
