Rule-adhering synthetic data – the lingua franca of learning

09/12/2022
by   Michael Platzer, et al.
0

AI-generated synthetic data allows to distill the general patterns of existing data, that can then be shared safely as granular-level representative, yet novel data samples within the original semantics. In this work we explore approaches of incorporating domain expertise into the data synthesis, to have the statistical properties as well as pre-existing domain knowledge of rules be represented. The resulting synthetic data generator, that can be probed for any number of new samples, can then serve as a common source of intelligence, as a lingua franca of learning, consumable by humans and machines alike. We demonstrate the concept for a publicly available data set, and evaluate its benefits via descriptive analysis as well as a downstream ML model.

READ FULL TEXT

page 2

page 3

research
04/07/2021

Representative Fair Synthetic Data

Algorithms learn rules and associations based on the training data that ...
research
05/26/2023

On Consistent Bayesian Inference from Synthetic Data

Generating synthetic data, with or without differential privacy, has att...
research
11/19/2022

An experimental study on Synthetic Tabular Data Evaluation

In this paper, we present the findings of various methodologies for meas...
research
01/10/2022

A statistical shape model for radiation-free assessment and classification of craniosynostosis

The assessment of craniofacial deformities requires patient data which i...
research
05/09/2023

Novel Synthetic Data Tool for Data-Driven Cardboard Box Localization

Application of neural networks in industrial settings, such as automated...
research
03/29/2021

A Model-Based Approach to Synthetic Data Set Generation for Patient-Ventilator Waveforms for Machine Learning and Educational Use

Although mechanical ventilation is a lifesaving intervention in the ICU,...
research
05/06/2020

Shape of synth to come: Why we should use synthetic data for English surface realization

The Surface Realization Shared Tasks of 2018 and 2019 were Natural Langu...

Please sign up or login with your details

Forgot password? Click here to reset