Estimating related words computationally using language model from the Mahabharata – an Indian epic

05/09/2023
by   Vrunda Gadesha, et al.
0

'Mahabharata' is the most popular among many Indian pieces of literature referred to in many domains for completely different purposes. This text itself is having various dimension and aspects which is useful for the human being in their personal life and professional life. This Indian Epic is originally written in the Sanskrit Language. Now in the era of Natural Language Processing, Artificial Intelligence, Machine Learning, and Human-Computer interaction this text can be processed according to the domain requirement. It is interesting to process this text and get useful insights from Mahabharata. The limitation of the humans while analyzing Mahabharata is that they always have a sentiment aspect towards the story narrated by the author. Apart from that, the human cannot memorize statistical or computational details, like which two words are frequently coming in one sentence? What is the average length of the sentences across the whole literature? Which word is the most popular word across the text, what are the lemmas of the words used across the sentences? Thus, in this paper, we propose an NLP pipeline to get some statistical and computational insights along with the most relevant word searching method from the largest epic 'Mahabharata'. We stacked the different text-processing approaches to articulate the best results which can be further used in the various domain where Mahabharata needs to be referred.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
02/05/2023

A Semantic Approach to Negation Detection and Word Disambiguation with Natural Language Processing

This study aims to demonstrate the methods for detecting negations in a ...
research
07/21/2020

Human Abnormality Detection Based on Bengali Text

In the field of natural language processing and human-computer interacti...
research
08/13/2021

Generalized Optimal Linear Orders

The sequential structure of language, and the order of words in a senten...
research
03/04/2023

Variational Quantum Classifiers for Natural-Language Text

As part of the recent research effort on quantum natural language proces...
research
08/24/2023

Separating the Human Touch from AI-Generated Text using Higher Criticism: An Information-Theoretic Approach

We propose a method to determine whether a given article was entirely wr...
research
12/19/2019

Identifying Adversarial Sentences by Analyzing Text Complexity

Attackers create adversarial text to deceive both human perception and t...
research
02/15/2019

Contextual Word Representations: A Contextual Introduction

This introduction aims to tell the story of how we put words into comput...

Please sign up or login with your details

Forgot password? Click here to reset