Ordered Sets for Data Analysis

08/27/2019
by   Sergei O. Kuznetsov, et al.
0

This book dwells on mathematical and algorithmic issues of data analysis based on generality order of descriptions and respective precision. To speak of these topics correctly, we have to go some way getting acquainted with the important notions of relation and order theory. On the one hand, data often have a complex structure with natural order on it. On the other hand, many symbolic methods of data analysis and machine learning allow to compare the obtained classifiers w.r.t. their generality, which is also an order relation. Efficient algorithms are very important in data analysis, especially when one deals with big data, so scalability is a real issue. That is why we analyze the computational complexity of algorithms and problems of data analysis. We start from the basic definitions and facts of algorithmic complexity theory and analyze the complexity of various tools of data analysis we consider. The tools and methods of data analysis, like computing taxonomies, groups of similar objects (concepts and n-clusters), dependencies in data, classification, etc., are illustrated with applications in particular subject domains, from chemoinformatics to text mining and natural language processing.

READ FULL TEXT

Authors

page 1

page 2

page 3

page 4

06/03/2019

An Introduction to a New Text Classification and Visualization for Natural Language Processing Using Topological Data Analysis

Topological Data Analysis (TDA) is a novel new and fast growing field of...
05/20/2019

Tools for analyzing R code the tidy way

With the current emphasis on reproducibility and replicability, there is...
04/25/2021

Breiman's two cultures: You don't have to choose sides

Breiman's classic paper casts data analysis as a choice between two cult...
05/30/2017

The Role of Data Analysis in the Development of Intelligent Energy Networks

Data analysis plays an important role in the development of intelligent ...
10/14/2019

code::proof: Prepare for most weather conditions

Computational tools for data analysis are being released daily on reposi...
10/08/2010

Algorithmic and Statistical Perspectives on Large-Scale Data Analysis

In recent years, ideas from statistics and scientific computing have beg...
07/21/2019

Logical Classification of Partially Ordered Data

Issues concerning intelligent data analysis occurring in machine learnin...
This week in AI

Get the week's most popular data science and artificial intelligence research sent straight to your inbox every Saturday.