Barbara Plank
4 papers in the PaperMetrix corpus
Papers by this author
-
DaN+: Danish Nested Named Entities and Lexical Normalization
2020
This paper introduces DAN+, a new multi-domain corpus and annotation guidelines for Danish nested named entities (NEs) and lexical normalization to support research on cross-lingual cross-domain learning for a less-resourced language. We empirically assess three …
-
Genre as Weak Supervision for Cross-lingual Dependency Parsing
2021 · Proceedings of the 2021 Conference on Empirical Methods in Natural Language Processing
Recent work has shown that monolingual masked language models learn to represent data-driven notions of language variation which can be used for domain-targeted training data selection. Dataset genre labels are already frequently available, yet remain …
-
A Rose by Any Other Name: LLM-Generated Explanations Are Good Proxies for Human Explanations to Collect Label Distributions on NLI
2025
Disagreement in human labeling is ubiquitous, and can be captured in human judgment distributions (HJDs).Recent research has shown that explanations provide valuable information for understanding human label variation (HLV) and large language models (LLMs) can …
-
References Matter: Investigating the Impact of Reference Set Variation on Summarization Evaluation
2025 · arXiv (Cornell University)
Human language production exhibits remarkable richness and variation, reflecting diverse communication styles and intents. However, this variation is often overlooked in summarization evaluation. While having multiple reference summaries is known to improve correlation with human …