Search papers, labs, and topics across Lattice.
This paper establishes new exponential-in-dimension lower bounds for the Maximum Halfspace Discrepancy problem, a model for linear classification, by reducing from hardness conjectures for Affine Degeneracy testing and $k$-Sum problems. The reductions yield near-matching lower bounds of $\tilde\Omega(n^d)$ and $\tilde\Omega(1/\varepsilon^d)$ based on Affine Degeneracy testing, and $\tilde\Omega(n^{d/2})$ and $\tilde\Omega(1/\varepsilon^{d/2})$ conditioned on $k$-Sum. Notably, the $\tilde\Omega(n^d)$ bound holds unconditionally when restricting the computational model to sidedness queries, a common setting in many algorithms.
Linear classification, a cornerstone of machine learning, is provably harder than we thought in high dimensions.
We establish new exponential in dimension lower bounds for the Maximum Halfspace Discrepancy problem, which models linear classification. Both are fundamental problems in computational geometry and machine learning in their exact and approximate forms. However, only $O(n^d)$ and respectively $\tilde O(1/\varepsilon^d)$ upper bounds are known and complemented by polynomial lower bounds that do not support the exponential in dimension dependence. We close this gap up to polylogarithmic terms by reduction from widely-believed hardness conjectures for Affine Degeneracy testing and $k$-Sum problems. Our reductions yield matching lower bounds of $\tilde\Omega(n^d)$ and respectively $\tilde\Omega(1/\varepsilon^d)$ based on Affine Degeneracy testing, and $\tilde\Omega(n^{d/2})$ and respectively $\tilde\Omega(1/\varepsilon^{d/2})$ conditioned on $k$-Sum. The first bound also holds unconditionally if the computational model is restricted to make sidedness queries, which corresponds to a widely spread setting implemented and optimized in many contemporary algorithms and computing paradigms.