Pepa Atanasova
4 papers in the PaperMetrix corpus
Papers by this author
-
A Large-Scale Semi-Supervised Dataset for Offensive Language Identification
2020 · arXiv (Cornell University)
The use of offensive language is a major problem in social media which has led to an abundance of research in detecting content such as hate speech, cyberbulling, and cyber-aggression. There have been several attempts …
-
SOLID: A Large-Scale Semi-Supervised Dataset for Offensive Language Identification
2021
The widespread use of offensive content in social media has led to an abundance of research in detecting language such as hate speech, cyberbullying, and cyber-aggression. Recent work presented the OLID dataset, which follows a …
-
Diagnostics-Guided Explanation Generation
2021 · arXiv (Cornell University)
Explanations shed light on a machine learning model's rationales and can aid in identifying deficiencies in its reasoning process. Explanation generation models are typically trained in a supervised way given human explanations. When such annotations …
-
Explaining Interactions Between Text Spans
2023 · arXiv (Cornell University)
Reasoning over spans of tokens from different parts of the input is essential for natural language understanding (NLU) tasks such as fact-checking (FC), machine reading comprehension (MRC) or natural language inference (NLI). However, existing highlight-based …