preprint Open access

Overview for the Second Shared Task on Language Identification in Code-Switched Data

  • arXiv (Cornell University)
  • Cornell University
Research footprint

At a glance

Citations
29
References
12
Comments
0
Paper overview

Abstract

We present an overview of the second shared task on language identification in code-switched data. For the shared task, we had code-switched data from two different language pairs: Modern Standard Arabic-Dialectal Arabic (MSA-DA) and Spanish-English (SPA-ENG). We had a total of nine participating teams, with all teams submitting a system for SPA-ENG and four submitting for MSA-DA. Through evaluation, we found that once again language identification is more difficult for the language pair that is more closely related. We also found that this year's systems performed better overall than the systems from the previous shared task indicating overall progress in the state of the art for this task.

Record transparency

Publication details

DOI
10.48550/arxiv.1909.13016
OpenAlex
W2975529437
Document type
preprint
Language
EN
Source
arXiv (Cornell University)
Last metadata update
Community

Comments

Log in to join the discussion.

  1. No comments yet. Start the discussion.