A Novel Approach to Enhance the Performance of Semantic Search in Bengali using Neural Net and other Classification Techniques

11/04/2019
by   Arijit Das, et al.
0

Search has for a long time been an important tool for users to retrieve information. Syntactic search is matching documents or objects containing specific keywords like user-history, location, preference etc. to improve the results. However, it is often possible that the query and the best answer have no term or very less number of terms in common and syntactic search can not perform properly in such cases. Semantic search, on the other hand, resolves these issues but suffers from lack of annotation, absence of WordNet in case of low resource languages. In this work, we have demonstrated an end to end procedure to improve the performance of semantic search using semi-supervised and unsupervised learning algorithms. An available Bengali repository was chosen to have seven types of semantic properties primarily to develop the system. Performance has been tested using Support Vector Machine, Naive Bayes, Decision Tree and Artificial Neural Network (ANN). Our system has achieved the efficiency to predict the correct semantics using knowledge base over the time of learning. A repository containing around a million sentences, a product of TDIL project of Govt. of India, was used to test our system at first instance. Then the testing has been done for other languages. Being a cognitive system it may be very useful for improving user satisfaction in e-Governance or m-Governance in the multilingual environment and also for other applications.

READ FULL TEXT
research
11/24/2020

Acoustic span embeddings for multilingual query-by-example search

Query-by-example (QbE) speech search is the task of matching spoken quer...
research
01/27/2021

Automatic image annotation base on Naive Bayes and Decision Tree classifiers using MPEG-7

Recently it has become essential to search for and retrieve high-resolut...
research
07/29/2018

Discovering Latent Information By Spreading Activation Algorithm For Document Retrieval

Syntactic search relies on keywords contained in a query to find suitabl...
research
05/21/2022

Supplementary Results of a Comparative Syntactic and Semantic Study of Terms for Software Testing Glossaries

This preprint specifies supplementary material and the results of a comp...
research
09/15/2023

Multilingual Sentence-Level Semantic Search using Meta-Distillation Learning

Multilingual semantic search is the task of retrieving relevant contents...
research
11/19/2019

Neural Network based End-to-End Query by Example Spoken Term Detection

This paper focuses on the problem of query by example spoken term detect...
research
12/08/2021

ADBCMM : Acronym Disambiguation by Building Counterfactuals and Multilingual Mixing

Scientific documents often contain a large number of acronyms. Disambigu...

Please sign up or login with your details

Forgot password? Click here to reset