Hamdy Mubarak
6 papers in the PaperMetrix corpus
Papers by this author
-
Arabic Dialect Identification in the Wild
2020 · arXiv (Cornell University)
We present QADI, an automatically collected dataset of tweets belonging to a wide range of country-level Arabic dialects -covering 18 different countries in the Middle East and North Africa region. Our method for building this …
-
Fact-Checking the Output of Large Language Models via Token-Level Uncertainty Quantification
2024 · arXiv (Cornell University)
Large language models (LLMs) are notorious for hallucinating, i.e., producing erroneous claims in their output. Such hallucinations can be dangerous, as occasional factual inaccuracies in the generated text might be obscured by the rest of …
-
SemEval-2016 Task 3: Community Question Answering
2016
Preslav Nakov, Lluís Màrquez, Alessandro Moschitti, Walid Magdy, Hamdy Mubarak, Abed Alhakim Freihat, Jim Glass, Bilal Randeree. Proceedings of the 10th International Workshop on Semantic Evaluation (SemEval-2016). 2016.
-
Farasa: A Fast and Furious Segmenter for Arabic
2016
In this paper, we present Farasa, a fast and accurate Arabic segmenter. Our approach is based on SVM-rank using linear kernels. We measure the performance of the segmenter in terms of accuracy and efficiency, in …
-
Farasa: A New Fast and Accurate Arabic Word Segmenter
2016
In this paper, we present Farasa (meaning insight in Arabic), which is a fast and accurate Arabic segmenter.Segmentation involves breaking Arabic words into their constituent clitics.Our approach is based on SVM rank using linear kernels.The …
-
SemEval-2017 Task 3: Community Question Answering
2017
Preslav Nakov, Doris Hoogeveen, Lluís Màrquez, Alessandro Moschitti, Hamdy Mubarak, Timothy Baldwin, Karin Verspoor. Proceedings of the 11th International Workshop on Semantic Evaluation (SemEval-2017). 2017.