arXiv:2409.13453 Abstract | arXiv Analytics

arXiv:2409.13453 [math.NA]Abstract References Reviews Resources

Data Compression using Rank-1 Lattices for Parameter Estimation in Machine Learning

Michael Gnewuch, Kumar Harsha, Marcin Wnuk

Published 2024-09-20Version 1

The mean squared error and regularized versions of it are standard loss functions in supervised machine learning. However, calculating these losses for large data sets can be computationally demanding. Modifying an approach of J. Dick and M. Feischl [Journal of Complexity 67 (2021)], we present algorithms to reduce extensive data sets to a smaller size using rank-1 lattices. Rank-1 lattices are quasi-Monte Carlo (QMC) point sets that are, if carefully chosen, well-distributed in a multidimensional unit cube. The compression strategy in the preprocessing step assigns every lattice point a pair of weights depending on the original data and responses, representing its relative importance. As a result, the compressed data makes iterative loss calculations in optimization steps much faster. We analyze the errors of our QMC data compression algorithms and the cost of the preprocessing step for functions whose Fourier coefficients decay sufficiently fast so that they lie in certain Wiener algebras or Korobov spaces. In particular, we prove that our approach can lead to arbitrary high convergence rates as long as the functions are sufficiently smooth.

Comments: 25 pages, 1 figure

Categories: math.NA, cs.LG, cs.NA, stat.ML

Subjects: 68Q32, 65D30, 42B05, 11K38, F.2.1, G.1.2

Keywords: parameter estimation, machine learning, arbitrary high convergence rates, fourier coefficients decay sufficiently fast, qmc data compression algorithms

Related articles: Most relevant | Search more

arXiv:2009.14596 [math.NA] (Published 2020-09-23)

Machine Learning and Computational Mathematics

Weinan E

arXiv:2009.02687 [math.NA] (Published 2020-09-06)

Nonlinear reduced models for state and parameter estimation

Albert Cohen, Wolfgang Dahmen, Olga Mula, James Nichols

arXiv:2410.12654 [math.NA] (Published 2024-10-16)

A comparative analysis of metamodels for lumped cardiovascular models, and pipeline for sensitivity analysis, parameter estimation, and uncertainty quantification

John M. Hanna, Pavlos Varsos, Jérôme Kowalski, Lorenzo Sala, Roel Meiburg, Irene E. Vignon-Clementel

arXiv Analytics

arXiv:2409.13453 [math.NA]Abstract References Reviews Resources

Data Compression using Rank-1 Lattices for Parameter Estimation in Machine Learning

Links

Toolbox

arXiv:2409.13453 [math.NA]AbstractReferencesReviewsResources

Data Compression using Rank-1 Lattices for Parameter Estimation in Machine Learning

Links

Toolbox

arXiv:2409.13453 [math.NA]Abstract References Reviews Resources