Akshat Shrivastava
3 papers in the PaperMetrix corpus
Papers by this author
-
Muppet: Massive Multi-task Representations with Pre-Finetuning
2021 · arXiv (Cornell University)
We propose pre-finetuning, an additional large-scale learning stage between language model pre-training and fine-tuning. Pre-finetuning is massively multi-task learning (around 50 datasets, over 4.8 million total labeled examples), and is designed to encourage learning of …
-
CoDi: Conversational Distillation for Grounded Question Answering
2024 · arXiv (Cornell University)
Distilling conversational skills into Small Language Models (SLMs) with approximately 1 billion parameters presents significant challenges. Firstly, SLMs have limited capacity in their model parameters to learn extensive knowledge compared to larger models. Secondly, high-quality …
-
PRoDeliberation: Parallel Robust Deliberation for End-to-End Spoken Language Understanding
2024
Spoken Language Understanding (SLU) is a critical component of voice assistants; it consists of converting speech to semantic parses for task execution.Previous works have explored end-to-end models to improve the quality and robustness of SLU …