Search papers, labs, and topics across Lattice.
This study investigates the impact of additive label noise on spectral learning methods used for supervised regression, revealing that noise leads to a predictable drift in the learned coefficient vector. By deriving a closed-form expression for the overlap between noisy and noiseless coefficient vectors after whitening the empirical feature geometry, the authors establish a universal degradation curve dictated by an intrinsic noise scale. Numerical experiments across various bases confirm that there is a fundamental noise threshold that destabilizes coefficient estimates, highlighting intrinsic limitations in recovering functional relationships from noisy data.
Spectral learning faces a fundamental noise threshold that can destabilize coefficient estimates, challenging the reliability of functional recovery from noisy data.
Learning functional relationships from noisy data is a central problem in scientific inference. Spectral methods approximate unknown functions by expanding them in a basis and estimating the corresponding coefficients from data, but the stability of these coefficients under noise remains poorly understood. Here we study supervised regression with additive label noise using sparse spectral representations across multiple bases and dimensions. We show that noise induces a predictable drift in the learned coefficient vector whose magnitude depends on the effective number of active spectral modes. After whitening the empirical feature geometry, we derive a closed-form expression for the overlap between noisy and noiseless coefficient vectors, revealing a universal degradation curve governed by a single intrinsic noise scale. Numerical experiments across Fourier, Legendre, Bessel, and Haar bases confirm the theoretical prediction. The results demonstrate that spectral learning exhibits a fundamental noise threshold beyond which coefficient estimates become unstable, placing intrinsic limits on recovering functional structure from noisy data.