login

Naive (Bayes) at forty: The independence assumption in information retrieval

Lecture notes in computer sciencePublished 1 January 1998Open access
David Lewis
Citations2,092
SJR quartileQ2
SJR score0.35
SNIP0.55
View PDF

TL;DR

The naive Bayes classifier, currently experiencing a renaissance in machine learning, has long been a core technique in information retrieval, and some of the variations used for text retrieval and classification are reviewed.

Abstract

The naive Bayes classifier, currently experiencing a renaissance ] in machine learning, has long been a core technique in information retrieval. We review some of the variations of naive Bayes models used for text retrieval and classification, focusing on the distributional assumptions made about word occurrences in documents.

Keywords

Computer Science