Subhabrata Mukherjee
5 papers in the PaperMetrix corpus
Papers by this author
-
TinyMBERT: Multi-Stage Distillation Framework for Massive Multi-lingual NER.
2020 · arXiv (Cornell University)
Deep and large pre-trained language models are the state-of-the-art for various natural language processing tasks. However, the huge size of these models could be a deterrent to use them in practice. Some recent and concurrent …
-
Product Insights: Analyzing Product Intents in Web Search
2020 · arXiv (Cornell University)
Web search engines are frequently used to access information about products. This has increased in recent times with the rising popularity of e-commerce. However, there is limited understanding of what users search for and their …
-
RED QUEEN: Safeguarding Large Language Models against Concealed Multi-Turn Jailbreaking
2024 · arXiv (Cornell University)
The rapid progress of Large Language Models (LLMs) has opened up new opportunities across various domains and applications; yet it also presents challenges related to potential misuse. To mitigate such risks, red teaming has been …
-
OpenTag
2018
Extraction of missing attribute values is to find values describing an attribute of interest from a free text input. Most past related work on extraction of missing attribute values work with a closed world assumption …
-
Gender Bias in Multilingual Embeddings and Cross-Lingual Transfer
2020
Multilingual representations embed words from many languages into a single semantic space such that words with similar meanings are close to each other regardless of the language. These embeddings have been widely used in various …