Automatic Analysis of Linguistic Features in Journal Articles of Different Academic Impacts with Feature Engineering Techniques

11/15/2021
by   Siyu Lei, et al.
0

English research articles (RAs) are an essential genre in academia, so the attempts to employ NLP to assist the development of academic writing ability have received considerable attention in the last two decades. However, there has been no study employing feature engineering techniques to investigate the linguistic features of RAs of different academic impacts (i.e., the papers of high/moderate citation times published in the journals of high/moderate impact factors). This study attempts to extract micro-level linguistic features in high- and moderate-impact journal RAs, using feature engineering methods. We extracted 25 highly relevant features from the Corpus of English Journal Articles through feature selection methods. All papers in the corpus deal with COVID-19 medical empirical studies. The selected features were then validated of the classification performance in terms of consistency and accuracy through supervised machine learning methods. Results showed that 24 linguistic features such as the overlapping of content words between adjacent sentences, the use of third-person pronouns, auxiliary verbs, tense, emotional words provide consistent and accurate predictions for journal articles with different academic impacts. Lastly, the random forest model is shown to be the best model to fit the relationship between these 24 features and journal articles with high and moderate impacts. These findings can be used to inform academic writing courses and lay the foundation for developing automatic evaluation systems for L2 graduate students.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
11/23/2017

Microsoft Academic Automatic Document Searches: Accuracy for Journal Articles and Suitability for Citation Analysis

Microsoft Academic is a free academic search engine and citation index t...
research
07/13/2021

What do writing features tell us about AI papers?

As the numbers of submissions to conferences grow quickly, the task of a...
research
01/27/2020

Determining crucial factors for the popularity of scientific articles

Using a set of over 70.000 records from PLOS One journal consisting of 3...
research
07/25/2022

Analysis of the deletions of DOIs: What factors undermine their persistence and to what extent?

Digital Object Identifiers (DOIs) are regarded as persistent; however, t...
research
05/07/2019

Does Environmental Economics lead to patentable research?

In this feasibility study, the impact of academic research from social s...
research
11/28/2021

Enhancing Identification of Structure Function of Academic Articles Using Contextual Information

With the enrichment of literature resources, researchers are facing the ...
research
10/06/2022

A Machine Learning Based Approach to Categorize Research Journals

In this modern technological era, categorization and ranking of research...

Please sign up or login with your details

Forgot password? Click here to reset