preprint
وصول مفتوح
Improvement in Machine Translation from English to Punjabi by Identifying the Morpheme Boundaries
Research footprint
At a glance
- الاستشهادات
- 0
- المراجع
- 0
- Comments
- 0
Paper overview
Abstract
We discuss about language distinct and an unsupervised approach for Morphological Analysis. The algorithm is based on probability that uses the distance, frequency and length of the strings. In future it would solve problems of large corpora and agglutinative languages as well. We perform the algorithm on English data as well as Punjabi data and get the results as follows, as the number of morphemes recognized are more in English than in Punjabi language due to the fluctuations in random behavior showing smaller segmentations in small data sizes. There will be always a room for change ahead as the language grows.
Record transparency
Publication details
- OpenAlex
- W3159605918
- Document type
- preprint
- Language
- EN
- Source
- SSRN Electronic Journal
- Last metadata update
Comments
تسجيل الدخول للانضمام إلى النقاش.