Deep Learning-based approaches for automatic detection of shell nouns and evaluation on WikiText-2

08/25/2022
by   Chengdong Yao, et al.
0

In some areas, such as Cognitive Linguistics, researchers are still using traditional techniques based on manual rules and patterns. Since the definition of shell noun is rather subjective and there are many exceptions, this time-consuming work had to be done by hand in the past when Deep Learning techniques were not mature enough. With the increasing number of networked languages, these rules are becoming less useful. However, there is a better alternative now. With the development of Deep Learning, pre-trained language models have provided a good technical basis for Natural Language Processing. Automated processes based on Deep Learning approaches are more in line with modern needs. This paper collaborates across borders to propose two Neural Network models for the automatic detection of shell nouns and experiment on the WikiText-2 dataset. The proposed approaches not only allow the entire process to be automated, but the precision has reached 94 articles, comparable to that of human annotators. This shows that the performance and generalization ability of the model is good enough to be used for research purposes. Many new nouns are found that fit the definition of shell noun very well. All discovered shell nouns as well as pre-trained models and code are available on GitHub.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
07/02/2018

Make (Nearly) Every Neural Network Better: Generating Neural Network Ensembles by Weight Parameter Resampling

Deep Neural Networks (DNNs) have become increasingly popular in computer...
research
02/21/2021

Automatic Code Generation using Pre-Trained Language Models

Recent advancements in natural language processing <cit.> <cit.> have le...
research
02/11/2022

Similarity learning for wells based on logging data

One of the first steps during the investigation of geological objects is...
research
01/29/2023

Boosting Automated Patch Correctness Prediction via Pre-trained Language Model

Automated program repair (APR) aims to fix software bugs automatically w...
research
09/29/2022

FastPacket: Towards Pre-trained Packets Embedding based on FastText for next-generation NIDS

New Attacks are increasingly used by attackers everyday but many of them...
research
01/12/2023

Inaccessible Neural Language Models Could Reinvigorate Linguistic Nativism

Large Language Models (LLMs) have been making big waves in the machine l...
research
10/04/2021

The state-of-the-art in text-based automatic personality prediction

Personality detection is an old topic in psychology and Automatic Person...

Please sign up or login with your details

Forgot password? Click here to reset