article Open access

Intelligent text classification system based on self-administered ontology

  • TURKISH JOURNAL OF ELECTRICAL ENGINEERING & COMPUTER SCIENCES
  • Scientific and Technological Research Council of Turkey (TUBITAK)
Research footprint

At a glance

Citations
5
References
23
Comments
0
Paper overview

Abstract

Over the last couple of decades, web classification has gradually transitioned from a syntax- to semantic-centered approach that classifies the text based on domain ontologies. These ontologies are either built manually or populated automatically using machine learning techniques. A prerequisite condition to build such systems is the availability of ontology, which may be either full-fledged domain ontology or a seed ontology that can be enriched automatically. This is a dependency condition for any given semantics-based text classification system. We share the details of a proof of concept of a web classification system that is self-governed in terms of ontology population and does not require any prebuilt ontology, neither full-fledged nor seed. It starts from a user query, builds a seed ontology from it, and automatically enriches it by extracting concepts from the downloaded documents only. The evaluated parameters like precision (85{\%}), accuracy (86{\%}), AUC (convex), and MCC (high positive) demonstrate the better performance of the proposed system when compared with similar automated text classification systems.

Record transparency

Publication details

DOI
10.3906/elk-1305-112
OpenAlex
W1656359313
Document type
article
Language
EN
Source
TURKISH JOURNAL OF ELECTRICAL ENGINEERING & COMPUTER SCIENCES
Last metadata update
Community

Comments

Log in to join the discussion.

  1. No comments yet. Start the discussion.