PCA, SVD, and Centering of Data

07/27/2023
by   Donggun Kim, et al.
0

The research detailed in this paper scrutinizes Principal Component Analysis (PCA), a seminal method employed in statistics and machine learning for the purpose of reducing data dimensionality. Singular Value Decomposition (SVD) is often employed as the primary means for computing PCA, a process that indispensably includes the step of centering - the subtraction of the mean location from the data set. In our study, we delve into a detailed exploration of the influence of this critical yet often ignored or downplayed data centering step. Our research meticulously investigates the conditions under which two PCA embeddings, one derived from SVD with centering and the other without, can be viewed as aligned. As part of this exploration, we analyze the relationship between the first singular vector and the mean direction, subsequently linking this observation to the congruity between two SVDs of centered and uncentered matrices. Furthermore, we explore the potential implications arising from the absence of centering in the context of performing PCA via SVD from a spectral analysis standpoint. Our investigation emphasizes the importance of a comprehensive understanding and acknowledgment of the subtleties involved in the computation of PCA. As such, we believe this paper offers a crucial contribution to the nuanced understanding of this foundational statistical method and stands as a valuable addition to the academic literature in the field of statistics.

READ FULL TEXT

page 6

page 7

page 8

page 9

research
06/24/2021

Regularisation for PCA- and SVD-type matrix factorisations

Singular Value Decomposition (SVD) and its close relative, Principal Com...
research
06/28/2019

FameSVD: Fast and Memory-efficient Singular Value Decomposition

We propose a novel algorithm to perform the Singular Value Decomposition...
research
02/08/2023

Mallat Scattering Transformation based surrogate for MagnetoHydroDynamics

A Machine and Deep Learning methodology is developed and applied to give...
research
10/19/2018

Heteroskedastic PCA: Algorithm, Optimality, and Applications

Principal component analysis (PCA) and singular value decomposition (SVD...
research
09/02/2020

A Survey of Singular Value Decomposition Methods for Distributed Tall/Skinny Data

The Singular Value Decomposition (SVD) is one of the most important matr...
research
07/15/2023

Identification of Stochasticity by Matrix-decomposition: Applied on Black Hole Data

Timeseries classification as stochastic (noise-like) or non-stochastic (...
research
03/25/2020

Derivation of Coupled PCA and SVD Learning Rules from a Newton Zero-Finding Framework

In coupled learning rules for PCA (principal component analysis) and SVD...

Please sign up or login with your details

Forgot password? Click here to reset