Discriminating between Similar Languages and Arabic Dialect Identification: A Report on the Third DSL Shared Task
At a glance
- Citations
- 167
- References
- 58
- Comments
- 0
Öz
We present the results of the third edition of the Discriminating between Similar Languages (DSL) shared task, which was organized as part of the VarDial’2016 workshop at COLING’2016. The challenge offered two subtasks: subtask 1 focused on the identification of very similar languages and language varieties in newswire texts, whereas subtask 2 dealt with Arabic dialect identification in speech transcripts. A total of 37 teams registered to participate in the task, 24 teams submitted test results, and 20 teams also wrote system description papers. High-order character n-grams were the most successful feature, and the best classification approaches included traditional supervised learning methods such as SVM, logistic regression, and language models, while deep learning approaches did not perform very well.
Publication details
- OpenAlex
- W2561747913
- Document type
- conference-paper
- Language
- EN
- Source
- International Conference on Computational Linguistics
- Last metadata update
Comments
Oturum Açın to join the discussion.