Condensed Representation of Machine Learning Data

12/29/2022
by   Rahman Salim Zengin, et al.
0

Training of a Machine Learning model requires sufficient data. The sufficiency of the data is not always about the quantity, but about the relevancy and reduced redundancy. Data-generating processes create massive amounts of data. When used raw, such big data is causing much computational resource utilization. Instead of using the raw data, a proper Condensed Representation can be used instead. Combining K-means, a well-known clustering method, with some correction and refinement facilities a novel Condensed Representation method for Machine Learning applications is introduced. To present the novel method meaningfully and visually, synthetically generated data is employed. It has been shown that by using the condensed representation, instead of the raw data, acceptably accurate model training is possible.

READ FULL TEXT

page 3

page 4

page 7

research
05/19/2021

Mill.jl and JsonGrinder.jl: automated differentiable feature extraction for learning from raw JSON data

Learning from raw data input, thus limiting the need for manual feature ...
research
12/13/2021

Data Collection and Quality Challenges in Deep Learning: A Data-Centric AI Perspective

Software 2.0 is a fundamental shift in software engineering where machin...
research
10/26/2020

The Representation Race - Preprocessing for Handling Time Phenomena

Designing the representation languages for the input, L E, and output, L...
research
08/01/2018

Towards Machine Learning on data from Professional Cyclists

Professional sports are developing towards increasingly scientific train...
research
09/08/2021

Understanding and Preparing Data of Industrial Processes for Machine Learning Applications

Industrial applications of machine learning face unique challenges due t...
research
09/23/2022

KeypartX: Graph-based Perception (Text) Representation

The availability of big data has opened up big opportunities for individ...
research
10/21/2020

How Can Businesses Best Leverage Data scrubbing?

The hype around data is hardly news. It is a critical part of a business...

Please sign up or login with your details

Forgot password? Click here to reset