conference-paper

Discriminating between Similar Languages and Arabic Dialect Identification: A Report on the Third DSL Shared Task

  • International Conference on Computational Linguistics
Research footprint

At a glance

Citations
167
References
58
Comments
0
Paper overview

Abstract

We present the results of the third edition of the Discriminating between Similar Languages (DSL) shared task, which was organized as part of the VarDial’2016 workshop at COLING’2016. The challenge offered two subtasks: subtask 1 focused on the identification of very similar languages and language varieties in newswire texts, whereas subtask 2 dealt with Arabic dialect identification in speech transcripts. A total of 37 teams registered to participate in the task, 24 teams submitted test results, and 20 teams also wrote system description papers. High-order character n-grams were the most successful feature, and the best classification approaches included traditional supervised learning methods such as SVM, logistic regression, and language models, while deep learning approaches did not perform very well.

Record transparency

Publication details

OpenAlex
W2561747913
Document type
conference-paper
Language
EN
Source
International Conference on Computational Linguistics
Last metadata update
Community

Comments

Log in to join the discussion.

  1. No comments yet. Start the discussion.