Data Augmentation for Mathematical Objects

07/13/2023
by   Tereso del Río, et al.
0

This paper discusses and evaluates ideas of data balancing and data augmentation in the context of mathematical objects: an important topic for both the symbolic computation and satisfiability checking communities, when they are making use of machine learning techniques to optimise their tools. We consider a dataset of non-linear polynomial problems and the problem of selecting a variable ordering for cylindrical algebraic decomposition to tackle these with. By swapping the variable names in already labelled problems, we generate new problem instances that do not require any further labelling when viewing the selection as a classification problem. We find this augmentation increases the accuracy of ML models by 63 this improvement is due to the balancing of the dataset and what is achieved thanks to further increasing the size of the dataset, concluding that both have a very significant effect. We finish the paper by reflecting on how this idea could be applied in other uses of machine learning in mathematics.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
04/24/2023

Explainable AI Insights for Symbolic Computation: A case study on selecting the variable ordering for cylindrical algebraic decomposition

In recent years there has been increased use of machine learning (ML) te...
research
05/22/2020

A machine learning based software pipeline to pick the variable ordering for algorithms with polynomial inputs

We are interested in the application of Machine Learning (ML) technology...
research
04/24/2019

Comparing machine learning models to choose the variable ordering for cylindrical algebraic decomposition

There has been recent interest in the use of machine learning (ML) appro...
research
06/03/2019

Algorithmically generating new algebraic features of polynomial systems for machine learning

There are a variety of choices to be made in both computer algebra syste...
research
05/25/2020

Towards a Robust WiFi-based Fall Detection with Adversarial Data Augmentation

Recent WiFi-based fall detection systems have drawn much attention due t...
research
02/28/2022

Realtime strategy for image data labelling using binary models and active sampling

Machine learning (ML) and Deep Learning (DL) tasks primarily depend on d...
research
09/29/2016

Experience with Heuristics, Benchmarks & Standards for Cylindrical Algebraic Decomposition

In the paper which inspired the SC-Square project, [E. Abraham, Building...

Please sign up or login with your details

Forgot password? Click here to reset