Researcher profile

Mona Diab

12 papers in the PaperMetrix corpus

Publications

Papers by this author

  1. The George Washington University System for the Code-Switching Workshop Shared Task 2016

    2016

    We describe our work in the EMNLP 2016 second code-switching shared task; a generic language independent framework for linguistic code switch point detection (LCSPD). The system uses characters level 5-grams and word level unigram language …

  2. Processing Dialectal Arabic: Exploiting Variability and Similarity to Overcome Challenges and Discover Opportunities

    2016 · International Conference on Computational Linguistics

    We recently witnessed an exponential growth in dialectal Arabic usage in both textual data and speech recordings especially in social media. Processing such media is of great utility for all kinds of applications ranging from …

  3. A Review on Language Models as Knowledge Bases

    2022 · arXiv (Cornell University)

    Recently, there has been a surge of interest in the NLP community on the use of pretrained Language Models (LMs) as Knowledge Bases (KBs). Researchers have shown that LMs trained on a sufficiently large (web) …

  4. Evaluating Multilingual Speech Translation under Realistic Conditions with Resegmentation and Terminology

    2023

    We present the ACL 60/60 evaluation sets for multilingual translation of ACL 2022 technical presentations into 10 target languages.This dataset enables further research into multilingual speech translation under realistic recording conditions with unsegmented audio and …

  5. SemEval-2015 Task 2: Semantic Textual Similarity, English, Spanish and Pilot on Interpretability

    2015

    Eneko Agirre, Carmen Banea, Claire Cardie, Daniel Cer, Mona Diab, Aitor Gonzalez-Agirre, Weiwei Guo, Iñigo Lopez-Gazpio, Montse Maritxalar, Rada Mihalcea, German Rigau, Larraitz Uria, Janyce Wiebe. Proceedings of the 9th International Workshop on Semantic Evaluation …

  6. Overview for the Second Shared Task on Language Identification in Code-Switched Data

    2016

    Giovanni Molina, Fahad AlGhamdi, Mahmoud Ghoneim, Abdelati Hawwari, Nicolas Rey-Villamizar, Mona Diab, Thamar Solorio. Proceedings of the Second Workshop on Computational Approaches to Code Switching. 2016.

  7. SemEval-2016 Task 1: Semantic Textual Similarity, Monolingual and Cross-Lingual Evaluation

    2016

    Eneko Agirre, Carmen Banea, Daniel Cer, Mona Diab, Aitor Gonzalez-Agirre, Rada Mihalcea, German Rigau, Janyce Wiebe. Proceedings of the 10th International Workshop on Semantic Evaluation (SemEval-2016). 2016.

  8. SemEval-2017 Task 1: Semantic Textual Similarity Multilingual and Crosslingual Focused Evaluation

    2017

    Semantic Textual Similarity (STS) measures the meaning similarity of sentences. Applications include machine translation (MT), summarization, generation, question answering (QA), short answer grading, semantic search, dialog and conversational systems. The STS shared task is a …

  9. Named Entity Recognition on Code-Switched Data: Overview of the CALCS 2018 Shared Task

    2018

    In the third shared task of the Computational Approaches to Linguistic Code-Switching (CALCS) workshop, we focus on Named Entity Recognition (NER) on code-switched social-media data. We divide the shared task into two competitions based on …

  10. Overview for the Second Shared Task on Language Identification in Code-Switched Data

    2019 · arXiv (Cornell University)

    We present an overview of the second shared task on language identification in code-switched data. For the shared task, we had code-switched data from two different language pairs: Modern Standard Arabic-Dialectal Arabic (MSA-DA) and Spanish-English …

  11. Analysing Off-The-Shelf Options for Question Answering with Portuguese FAQs

    2022 · DROPS (Schloss Dagstuhl – Leibniz Center for Informatics)

    Following the current interest in developing automatic question answering systems, we analyse alternative approaches for finding suitable answers from a list of Frequently Asked Questions (FAQs), in Portuguese. These rely on different technologies, some more …

  12. SemEval-2017 Task 1: Semantic Textual Similarity - Multilingual and Cross-lingual Focused Evaluation

    2017 · HAL (Le Centre pour la Communication Scientifique Directe)

    Semantic Textual Similarity (STS) measures the meaning similarity of sentences. Applications include machine translation (MT), summarization, generation, question answering (QA), short answer grading, semantic search, dialog and conversational systems. The STS shared task is a …