Ordered Sets for Data Analysis

08/27/2019
by   Sergei O. Kuznetsov, et al.
0

This book dwells on mathematical and algorithmic issues of data analysis based on generality order of descriptions and respective precision. To speak of these topics correctly, we have to go some way getting acquainted with the important notions of relation and order theory. On the one hand, data often have a complex structure with natural order on it. On the other hand, many symbolic methods of data analysis and machine learning allow to compare the obtained classifiers w.r.t. their generality, which is also an order relation. Efficient algorithms are very important in data analysis, especially when one deals with big data, so scalability is a real issue. That is why we analyze the computational complexity of algorithms and problems of data analysis. We start from the basic definitions and facts of algorithmic complexity theory and analyze the complexity of various tools of data analysis we consider. The tools and methods of data analysis, like computing taxonomies, groups of similar objects (concepts and n-clusters), dependencies in data, classification, etc., are illustrated with applications in particular subject domains, from chemoinformatics to text mining and natural language processing.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
06/03/2019

An Introduction to a New Text Classification and Visualization for Natural Language Processing Using Topological Data Analysis

Topological Data Analysis (TDA) is a novel new and fast growing field of...
research
03/25/2023

Exactly mergeable summaries

In the analysis of large/big data sets, aggregation (replacing values of...
research
05/20/2019

Tools for analyzing R code the tidy way

With the current emphasis on reproducibility and replicability, there is...
research
04/25/2021

Breiman's two cultures: You don't have to choose sides

Breiman's classic paper casts data analysis as a choice between two cult...
research
05/30/2017

The Role of Data Analysis in the Development of Intelligent Energy Networks

Data analysis plays an important role in the development of intelligent ...
research
05/19/2017

Foundations of Declarative Data Analysis Using Limit Datalog Programs

Motivated by applications in declarative data analysis, we study Datalog...
research
10/14/2019

code::proof: Prepare for most weather conditions

Computational tools for data analysis are being released daily on reposi...

Please sign up or login with your details

Forgot password? Click here to reset