Mandar Joshi
4 papers in the PaperMetrix corpus
Papers by this author
-
Realistic Evaluation Principles for Cross-document Coreference Resolution
2021 · arXiv (Cornell University)
We point out that common evaluation practices for cross-document coreference resolution have been unrealistically permissive in their assumed settings, yielding inflated results. We propose addressing this issue via two evaluation methodology principles. First, as in …
-
TriviaQA: A Large Scale Distantly Supervised Challenge Dataset for Reading Comprehension
2017
We present TriviaQA, a challenging reading comprehension dataset containing over 650K question-answer-evidence triples. TriviaQA includes 95K questionanswer pairs authored by trivia enthusiasts and independently gathered evidence documents, six per question on average, that provide high …
-
SpanBERT: Improving Pre-training by Representing and Predicting Spans
2020 · Transactions of the Association for Computational Linguistics
We present SpanBERT, a pre-training method that is designed to better represent and predict spans of text. Our approach extends BERT by (1) masking contiguous random spans, rather than random tokens, and (2) training the …
-
HISTORIAE, History of Socio-Cultural Transformation as Linguistic Data Science. A Humanities Use Case
2019 · DROPS (Schloss Dagstuhl – Leibniz Center for Informatics)
Given a combinatorial optimisation problem, there are typically multiple ways of modelling it for presentation to an automated solver. Choosing the right combination of model and target solver can have a significant impact on the …