Rank Is Not Capacity: Spectral Occupancy for Latent Graph Models

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Nakis, Nikolaos, Promponas, Panagiotis, Tsirkas, Konstantinos, Mamali, Katerina, Makri, Eftychia, Tassiulas, Leandros, Christakis, Nicholas A.
Format: Preprint
Published: 2026
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866911672235458560
author Nakis, Nikolaos
Promponas, Panagiotis
Tsirkas, Konstantinos
Mamali, Katerina
Makri, Eftychia
Tassiulas, Leandros
Christakis, Nicholas A.
author_facet Nakis, Nikolaos
Promponas, Panagiotis
Tsirkas, Konstantinos
Mamali, Katerina
Makri, Eftychia
Tassiulas, Leandros
Christakis, Nicholas A.
contents Graph representation learning has become a standard approach for analyzing networked data, with latent embeddings widely used for link prediction, community detection, and related tasks. Yet a basic design choice, the latent dimension, is still treated as a brittle hyperparameter, fixed before training and tuned by held-out performance. Learned factors are also identifiable only up to rotation and rescaling, so the nominal rank rarely coincides with the quantity that governs model behavior. We propose Spectral Prefix Extraction and Capacity-Targeted Representation Analysis (Spectra), which replaces rank as the unit of analysis with the spectrum of a learned positive semidefinite kernel, trace-normalized so that spectra are comparable across fits. The normalized eigenvalues form a distribution on the simplex, and their Shannon effective rank acts both as a summary of learned capacity and as a controllable training-time coordinate: a single scalar shapes this realized dimension during training, and bisection targets any desired value within the rank cap. To theoretically support that, we show local regularity and monotonicity of the realized-dimension profile. Across collaboration, social, biological, and infrastructure networks, Spectra traces performance--capacity frontiers that make the trade-off between predictive accuracy and realized dimension visible. It performs competitively with strong link-prediction baselines, yields aligned lower-capacity views of the same fitted model through spectral prefixes, and provides a principled handle on capacity in the overparameterized regime. Capacity thus becomes a property of the fitted model rather than a hyperparameter of the training.
format Preprint
id arxiv_https___arxiv_org_abs_2605_11142
institution arXiv
publishDate 2026
record_format arxiv
spellingShingle Rank Is Not Capacity: Spectral Occupancy for Latent Graph Models
Nakis, Nikolaos
Promponas, Panagiotis
Tsirkas, Konstantinos
Mamali, Katerina
Makri, Eftychia
Tassiulas, Leandros
Christakis, Nicholas A.
Machine Learning
Graph representation learning has become a standard approach for analyzing networked data, with latent embeddings widely used for link prediction, community detection, and related tasks. Yet a basic design choice, the latent dimension, is still treated as a brittle hyperparameter, fixed before training and tuned by held-out performance. Learned factors are also identifiable only up to rotation and rescaling, so the nominal rank rarely coincides with the quantity that governs model behavior. We propose Spectral Prefix Extraction and Capacity-Targeted Representation Analysis (Spectra), which replaces rank as the unit of analysis with the spectrum of a learned positive semidefinite kernel, trace-normalized so that spectra are comparable across fits. The normalized eigenvalues form a distribution on the simplex, and their Shannon effective rank acts both as a summary of learned capacity and as a controllable training-time coordinate: a single scalar shapes this realized dimension during training, and bisection targets any desired value within the rank cap. To theoretically support that, we show local regularity and monotonicity of the realized-dimension profile. Across collaboration, social, biological, and infrastructure networks, Spectra traces performance--capacity frontiers that make the trade-off between predictive accuracy and realized dimension visible. It performs competitively with strong link-prediction baselines, yields aligned lower-capacity views of the same fitted model through spectral prefixes, and provides a principled handle on capacity in the overparameterized regime. Capacity thus becomes a property of the fitted model rather than a hyperparameter of the training.
title Rank Is Not Capacity: Spectral Occupancy for Latent Graph Models
topic Machine Learning
url https://arxiv.org/abs/2605.11142