Marzena Karpinska
5 papers in the PaperMetrix corpus
Papers by this author
-
DEMETR: Diagnosing Evaluation Metrics for Translation
2022 · arXiv (Cornell University)
While machine translation evaluation metrics based on string overlap (e.g., BLEU) have their limitations, their computations are transparent: the BLEU score assigned to a particular candidate translation can be traced back to the presence or …
-
NarrativeTime: Dense Temporal Annotation on a Timeline
2019 · arXiv (Cornell University)
For the past decade, temporal annotation has been sparse: only a small portion of event pairs in a text was annotated. We present NarrativeTime, the first timeline-based annotation framework that achieves full coverage of all …
-
ezCoref: Towards Unifying Annotation Guidelines for Coreference Resolution
2022 · arXiv (Cornell University)
Large-scale, high-quality corpora are critical for advancing research in coreference resolution. However, existing datasets vary in their definition of coreferences and have been collected via complex and lengthy guidelines that are curated for linguistic experts. …
-
Error Span Annotation: A Balanced Approach for Human Evaluation of Machine Translation
2024 · arXiv (Cornell University)
High-quality Machine Translation (MT) evaluation relies heavily on human judgments. Comprehensive error classification methods, such as Multidimensional Quality Metrics (MQM), are expensive as they are time-consuming and can only be done by experts, whose availability …
-
Does quantization affect models' performance on long-context tasks?
2025 · arXiv (Cornell University)
Large language models (LLMs) now support context windows exceeding 128K tokens, but this comes with significant memory requirements and high inference latency. Quantization can mitigate these costs, but may degrade performance. In this work, we …