Researcher profile

Mitesh M. Khapra

7 papers in the PaperMetrix corpus

Publications

Papers by this author

  1. Diversity driven attention model for query-based abstractive summarization

    2017

    Abstractive summarization aims to generate a shorter version of the document covering all the salient points in a compact and coherent fashion. On the other hand, query-based summarization highlights those points that are relevant in …

  2. Let's Ask Again: Refine Network for Automatic Question Generation

    2019 · arXiv (Cornell University)

    In this work, we focus on the task of Automatic Question Generation (AQG) where given a passage and an answer the task is to generate the corresponding question. It is desired that the generated question …

  3. Bhasha-Abhijnaanam: Native-script and romanized Language Identification for 22 Indic languages

    2023 · arXiv (Cornell University)

    We create publicly available language identification (LID) datasets and models in all 22 Indian languages listed in the Indian constitution in both native-script and romanized text. First, we create Bhasha-Abhijnaanam, a language identification test set …

  4. Complex Sequential Question Answering: Towards Learning to Converse Over Linked Question Answer Pairs with a Knowledge Graph

    2018 · Proceedings of the AAAI Conference on Artificial Intelligence

    While conversing with chatbots, humans typically tend to ask many questions, a significant portion of which can be answered by referring to large-scale knowledge graphs (KG). While Question Answering (QA) and dialog systems have been …

  5. Towards a Better Metric for Evaluating Question Generation Systems

    2018

    There has always been criticism for using ngram based similarity metrics, such as BLEU, NIST, etc, for evaluating the performance of NLG systems. However, these metrics continue to remain popular and are recently being used …

  6. Towards Exploiting Background Knowledge for Building Conversation Systems

    2018

    Existing dialog datasets contain a sequence of utterances and responses without any explicit background knowledge associated with them. This has resulted in the development of models which treat conversation as a sequenceto-sequence generation task (i.e., …

  7. IndicNLPSuite: Monolingual Corpora, Evaluation Benchmarks and Pre-trained Multilingual Language Models for Indian Languages

    2020

    In this paper, we introduce NLP resources for 11 major Indian languages from two major language families. These resources include: (a) large-scale sentence-level monolingual corpora, (b) pre-trained word embeddings, (c) pre-trained language models, and (d) …