conference-paper

Intelligence Extraction Method of Domain Terms for Chinese Web Documents Based on Hierarchical Combination Strategy

Research footprint

At a glance

Citations
2
References
0
Comments
0
Paper overview

Abstract

Domain terms extraction based on Chinese Web document is an important step in the field of Chinese text information into machine recognition, and it is also the technical basis of intelligence information processing in domain ontology construction, text knowledge mining, and etc. The traditional methods of terms extracting are: Based on the dictionary, based on the rules and the statistical method. But each of methods contains some limitations, such as in single word processing, synonyms merging and so on. In order to solve these problems, the intelligence extraction method based on hierarchical combination strategy is proposed. It is divided into three layers: the first layer is the document preprocessing layer, the second layer is the words preparation layer, and the third layer is the term extraction layer. The effectiveness of this method is verified by experimental data in the field of weapon and equipment. Comparing with the traditional method, it is proved that this method has good accuracy and recall rate for the domain terms extraction based on the Chinese Web documents.

Record transparency

Publication details

DOI
10.1109/icitbs.2016.57
OpenAlex
W2758050689
Document type
conference-paper
Language
EN
Last metadata update
Community

Comments

Log in to join the discussion.

  1. No comments yet. Start the discussion.