conference-paper

Different Word Representation for Text Classification: A Comparative Study

Research footprint

At a glance

Citations
4
References
9
Comments
0
Paper overview

Abstract

Due to the large amounts of words usually present in documents, some of their appearances can complicate the classification process and make it less accurate. Accordingly, word representation methods have been employed to handle this issue through the use of a comparative study. In this study, we compare the effectiveness of both word embedding and TF-IDF weighting schema by applying four classifiers to assess the accuracy of the classification. To evaluate the effectiveness of our study, it was tested on the popular 20Newsgroup text document dataset. Following our experimentation, we found that using the TF-IDF method and ANN classifiers on the 20Newsgroup dataset greatly enhanced the text documents' classification compared against the use of word embedding and other classifiers.

Record transparency

Publication details

DOI
10.1109/aiccsa47632.2019.9035347
OpenAlex
W3010827580
Document type
conference-paper
Language
EN
Last metadata update
Community

Comments

Log in to join the discussion.

  1. No comments yet. Start the discussion.