conference-paper Open access

Language Resource Addition Strategies for Raw Text Parsing

Research footprint

At a glance

Citations
0
References
21
Comments
0
Paper overview

Öz

We focus on the improvement of accuracy of raw text parsing, from the viewpoint of language resource addition.In Japanese, the raw text parsing is divided into three steps: word segmentation, part-of-speech tagging, and dependency parsing.We investigate the contribution of language resource addition in each of three steps to the improvement in accuracy for two domain corpora.The experimental results show that this improvement depends on the target domain.For example, when we handle well-written texts of limited vocabulary, white paper, an effective language resource is a word-POS pair sequence corpus for the parsing accuracy.So we conclude that it is important to check out the characteristics of the target domain and to choose a suitable language resource addition strategy for the parsing accuracy improvement.

Record transparency

Publication details

DOI
10.63317/4im5mdmdkxyh
OpenAlex
W2576224118
Document type
conference-paper
Language
EN
Last metadata update
Community

Comments

Oturum Açın to join the discussion.

  1. No comments yet. Start the discussion.