Sharp concentration of uniform generalization errors in binary linear classification

Fuente: arXiv
Gespeichert in:
Bibliographische Detailangaben
1. Verfasser: Nakakita, Shogo
Format: Preprint
Veröffentlicht: 2025
Schlagworte:
Online-Zugang:
Tags: Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
_version_ 1866912451227811840
author Nakakita, Shogo
author_facet Nakakita, Shogo
contents We examine the concentration of uniform generalization errors around their expectation in binary linear classification problems via an isoperimetric argument. In particular, we establish Poincaré and log-Sobolev inequalities for the joint distribution of the output labels and the label-weighted input vectors, which we apply to derive concentration bounds. The derived concentration bounds are sharp up to moderate multiplicative constants by those under well-balanced labels. In asymptotic analysis, we also show that almost sure convergence of uniform generalization errors to their expectation occurs in very broad settings, such as proportionally high-dimensional regimes. Using this convergence, we establish uniform laws of large numbers under dimension-free conditions.
format Preprint
id arxiv_https___arxiv_org_abs_2505_16713
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle Sharp concentration of uniform generalization errors in binary linear classification
Nakakita, Shogo
Machine Learning
Statistics Theory
We examine the concentration of uniform generalization errors around their expectation in binary linear classification problems via an isoperimetric argument. In particular, we establish Poincaré and log-Sobolev inequalities for the joint distribution of the output labels and the label-weighted input vectors, which we apply to derive concentration bounds. The derived concentration bounds are sharp up to moderate multiplicative constants by those under well-balanced labels. In asymptotic analysis, we also show that almost sure convergence of uniform generalization errors to their expectation occurs in very broad settings, such as proportionally high-dimensional regimes. Using this convergence, we establish uniform laws of large numbers under dimension-free conditions.
title Sharp concentration of uniform generalization errors in binary linear classification
topic Machine Learning
Statistics Theory
url https://arxiv.org/abs/2505.16713