Leshem Choshen
5 papers in the PaperMetrix corpus
Papers by this author
-
Automatically Extracting Challenge Sets for Non-Local Phenomena in Neural Machine Translation
2019
We show that the state-of-the-art Transformer MT model is not biased towards monotonic reordering (unlike previous recurrent neural network models), but that nevertheless, longdistance dependencies remain a challenge for the model. Since most dependencies are …
-
Q2: : Evaluating Factual Consistency in Knowledge-Grounded Dialogues via Question Generation and Question Answering
2021 · Proceedings of the 2021 Conference on Empirical Methods in Natural Language Processing
Neural knowledge-grounded generative models for dialogue often produce content that is factually inconsistent with the knowledge they rely on, making them unreliable and limiting their applicability. Inspired by recent work on evaluating factual consistency in …
-
ComSum: Commit Messages Summarization and Meaning Preservation
2021 · arXiv (Cornell University)
We present ComSum, a data set of 7 million commit messages for text summarization. When documenting commits, software code changes, both a message and its summary are posted. We gather and filter those to curate …
-
The Grammar-Learning Trajectories of Neural Language Models
2022 · Proceedings of the 60th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers)
The learning trajectories of linguistic phenomena in humans provide insight into linguistic representation, beyond what can be gleaned from inspecting the behavior of an adult speaker. To apply a similar approach to analyze neural language …
-
tinyBenchmarks: evaluating LLMs with fewer examples
2024 · arXiv (Cornell University)
The versatility of large language models (LLMs) led to the creation of diverse benchmarks that thoroughly test a variety of language models' abilities. These benchmarks consist of tens of thousands of examples making evaluation of …