Search papers, labs, and topics across Lattice.
This paper investigates a continuum of degree-normalized spectral embeddings within the framework of random dot product graphs, establishing a row-wise central limit theorem that elucidates the impact of degree normalization on population geometry and local uncertainty of nodes. By employing a projected-Gaussian Bayes-error diagnostic, the authors compare various normalizations in two-community stochastic block models, revealing that the effectiveness of normalization is context-dependent, varying with network density and community structure. The findings highlight that stronger normalization tends to be advantageous in scenarios characterized by lower density or greater community imbalance, offering a nuanced understanding of spectral clustering methods.
No single normalization method excels universally in spectral clustering; effectiveness hinges on network density and community structure.
Spectral clustering methods for network data are commonly based on a few matrix representations, such as the adjacency matrix and the symmetric Laplacian. We study a continuum of degree-normalized spectral embeddings that includes these commonly used choices as special cases. Under a random dot product graph model, we establish a row-wise central limit theorem for this family of embeddings. The result provides an explicit description of how degree normalization affects both population geometry and the local uncertainty of embedded nodes. We use the limiting distributions to compare different normalizations in two-community stochastic block models through a projected-Gaussian Bayes-error diagnostic. These comparisons show that no single normalization is uniformly preferred. Instead, the favored normalization depends on network density, community imbalance, and block-probability structure. Typically, stronger normalization is favored in lower-density or more imbalanced settings. These results provide a unified distributional understanding of when and why alternative normalizations may improve spectral clustering.