Searching for chromate replacements using natural language processing and machine learning algorithms

08/11/2022
by   Shujing Zhao, et al.
0

The past few years has seen the application of machine learning utilised in the exploration of new materials. As in many fields of research - the vast majority of knowledge is published as text, which poses challenges in either a consolidated or statistical analysis across studies and reports. Such challenges include the inability to extract quantitative information, and in accessing the breadth of non-numerical information. To address this issue, the application of natural language processing (NLP) has been explored in several studies to date. In NLP, assignment of high-dimensional vectors, known as embeddings, to passages of text preserves the syntactic and semantic relationship between words. Embeddings rely on machine learning algorithms and in the present work, we have employed the Word2Vec model, previously explored by others, and the BERT model - applying them towards a unique challenge in materials engineering. That challenge is the search for chromate replacements in the field of corrosion protection. From a database of over 80 million records, a down-selection of 5990 papers focused on the topic of corrosion protection were examined using NLP. This study demonstrates it is possible to extract knowledge from the automated interpretation of the scientific literature and achieve expert human level insights.

READ FULL TEXT
research
04/17/2023

Use of social media and Natural Language Processing (NLP) in natural hazard research

Twitter is a microblogging service for sending short, public text messag...
research
07/22/2020

Multi-task learning for natural language processing in the 2020s: where are we going?

Multi-task learning (MTL) significantly pre-dates the deep learning era,...
research
11/15/2022

Searching for Carriers of the Diffuse Interstellar Bands Across Disciplines, using Natural Language Processing

The explosion of scientific publications overloads researchers with info...
research
09/01/2021

Latin writing styles analysis with Machine Learning: New approach to old questions

In the Middle Ages texts were learned by heart and spread using oral mea...
research
08/07/2018

Importance of the Mathematical Foundations of Machine Learning Methods for Scientific and Engineering Applications

There has been a lot of recent interest in adopting machine learning met...
research
06/02/2023

Analyzing Credit Risk Model Problems through NLP-Based Clustering and Machine Learning: Insights from Validation Reports

This paper explores the use of clustering methods and machine learning a...

Please sign up or login with your details

Forgot password? Click here to reset