conference-paper Open access

Generative Models For Indic Languages: Evaluating Content Generation Capabilities

Research footprint

At a glance

Citations
7
References
24
Comments
0
Paper overview

Abstract

Large language models (LLMs) and generative AI have emerged as the most important areas in the field of natural language processing (NLP).LLMs are considered to be a key component in several NLP tasks, such as summarization, question-answering, sentiment classification, and translation.Newer LLMs, such as Chat-GPT, BLOOMZ, and several such variants, are known to train on multilingual training data and hence are expected to process and generate text in multiple languages.Considering the widespread use of LLMs, evaluating their efficacy in multilingual settings is imperative.In this work, we evaluate the newest generative models (ChatGPT, mT0, and BLOOMZ) in the context of Indic languages.Specifically, we consider natural language generation (NLG) applications such as summarization and questionanswering in monolingual and cross-lingual settings.We observe that current generative models have limited capability for generating text in Indic languages in a zero-shot setting.In contrast, generative models perform consistently better on manual quality-based evaluation in Indic languages and English language generation.Considering limited generation performance, we argue that these LLMs are not intended to use in zero-shot fashion in downstream applications.

Record transparency

Publication details

DOI
10.26615/978-954-452-092-2_021
OpenAlex
W4388676221
Document type
conference-paper
Language
EN
Last metadata update
Community

Comments

Log in to join the discussion.

  1. No comments yet. Start the discussion.