Estimating related words computationally using language model from the Mahabharata -- an Indian epic

Gadesha, Vrunda; Joshi, Keyur D; Naik, Shefali

doi:10.1007/978-981-19-5224-1_63

Computer Science > Computation and Language

arXiv:2305.05420 (cs)

[Submitted on 9 May 2023]

Title:Estimating related words computationally using language model from the Mahabharata -- an Indian epic

Authors:Vrunda Gadesha, Keyur D Joshi, Shefali Naik

View PDF

Abstract:'Mahabharata' is the most popular among many Indian pieces of literature referred to in many domains for completely different purposes. This text itself is having various dimension and aspects which is useful for the human being in their personal life and professional life. This Indian Epic is originally written in the Sanskrit Language. Now in the era of Natural Language Processing, Artificial Intelligence, Machine Learning, and Human-Computer interaction this text can be processed according to the domain requirement. It is interesting to process this text and get useful insights from Mahabharata. The limitation of the humans while analyzing Mahabharata is that they always have a sentiment aspect towards the story narrated by the author. Apart from that, the human cannot memorize statistical or computational details, like which two words are frequently coming in one sentence? What is the average length of the sentences across the whole literature? Which word is the most popular word across the text, what are the lemmas of the words used across the sentences? Thus, in this paper, we propose an NLP pipeline to get some statistical and computational insights along with the most relevant word searching method from the largest epic 'Mahabharata'. We stacked the different text-processing approaches to articulate the best results which can be further used in the various domain where Mahabharata needs to be referred.

Comments:	ICT Analysis and Applications: Proceedings of ICT4SD 2022 pp 627-638
Subjects:	Computation and Language (cs.CL)
Cite as:	arXiv:2305.05420 [cs.CL]
	(or arXiv:2305.05420v1 [cs.CL] for this version)
	https://doi.org/10.48550/arXiv.2305.05420
Related DOI:	https://doi.org/10.1007/978-981-19-5224-1_63

Submission history

From: Vrunda Gadesha [view email]
[v1] Tue, 9 May 2023 13:13:26 UTC (966 KB)

Computer Science > Computation and Language

Title:Estimating related words computationally using language model from the Mahabharata -- an Indian epic

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computation and Language

Title:Estimating related words computationally using language model from the Mahabharata -- an Indian epic

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators