A performance analysis of a large language model for Marathi language NLP tasks.
Recently, Large Language Models (LLMs) have gained substantial attention due to their exceptional capabilities in various Natural Language Processing (NLP) tasks, particularly for widely spoken global languages. The increasing adoption of LLMs is primarily attributed to their ability to achieve near...
| Published in: | Language Resources & Evaluation Vol. 60; no. 3; pp. 1 - 37 |
|---|---|
| Main Authors: | , , |
| Format: | Article |
| Published: |
Springer Nature
Sep2026
|
| Online Access: | View this record in EBSCOhost |
| fields | @attributes: recordID: 1 pdfLink: plink: https://search.ebscohost.com/login.aspx?direct=true&db=hlh&AN=193845126&site=ehost-live header: @attributes: shortDbName: hlh uiTerm: 193845126 longDbName: Humanities International Complete uiTag: AN controlInfo: bkinfo: jinfo: jid: 1574020X 179V jtl: Language Resources & Evaluation issn: 1574020X maglogo: N pubinfo: dt: Sep2026 vid: 60 iid: 3 pid: 237 pub: Springer Nature artinfo: ui: 193845126 10.1007/s10579-026-09915-x ppf: 1 ppct: 36 formats: tig: atl: A performance analysis of a large language model for Marathi language NLP tasks. aug: au: Gaikwad, Harsha R. Laddha, Manjushree D. Kiwelekar, Arvind W. affil: https://ror.org/05tc4vc39 Department of Computer Engineering, Dr. Babasaheb Ambedkar Technological University, Lonere, Vidyavihar, 402 103, Maharashtra, Raigad, India sug: keyword: Communication and Culture Linguistics Language Large language model Large language model evaluation Marathi language tasks ab: Recently, Large Language Models (LLMs) have gained substantial attention due to their exceptional capabilities in various Natural Language Processing (NLP) tasks, particularly for widely spoken global languages. The increasing adoption of LLMs is primarily attributed to their ability to achieve near-human-level proficiency in language understanding and generation. However, the effectiveness of LLMs in regional languages requires a thorough evaluation before their deployment in NLP applications. This study conducts a performance analysis of OpenAI’s Generative Pre-trained Transformer (GPT) model specifically for Marathi, India’s third most widely spoken regional language. The research focuses on crucial NLP tasks, including Sentiment Analysis, Text Classification, and Paraphrase generation. This paper addresses the challenges of fine-tuning GPT model for regional languages and provides a detailed performance evaluation. The contributions of this study are twofold: firstly, it presents a diverse and validated dataset specifically designed for Marathi NLP tasks; secondly, it offers a detailed performance benchmarking of OpenAI’s GPT model in the context of paraphrasing, text classification, and sentiment analysis for the Marathi text. pubtype: Academic Journal doctype: Article src: R language: English refInfo: copyright: @attributes: flag: Y dt: @attributes: year: 2026 holdings: @attributes: islocal: N |
|---|