Hypothesis Spaces for Deep Learning

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Wang, Rui, Xu, Yuesheng, Yan, Mingsong
Format: Preprint
Published: 2024
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866913989307400192
author Wang, Rui
Xu, Yuesheng
Yan, Mingsong
author_facet Wang, Rui
Xu, Yuesheng
Yan, Mingsong
contents This paper introduces a hypothesis space for deep learning based on deep neural networks (DNNs). By treating a DNN as a function of two variables - the input variable and the parameter variable - we consider the set of DNNs where the parameter variable belongs to a space of weight matrices and biases determined by a prescribed depth and layer widths. To construct a Banach space of functions of the input variable, we take the weak* closure of the linear span of this DNN set. We prove that the resulting Banach space is a reproducing kernel Banach space (RKBS) and explicitly construct its reproducing kernel. Furthermore, we investigate two learning models - regularized learning and the minimum norm interpolation (MNI) problem - within the RKBS framework by establishing representer theorems. These theorems reveal that the solutions to these learning problems can be expressed as a finite sum of kernel expansions based on training data.
format Preprint
id arxiv_https___arxiv_org_abs_2403_03353
institution arXiv
publishDate 2024
record_format arxiv
spellingShingle Hypothesis Spaces for Deep Learning
Wang, Rui
Xu, Yuesheng
Yan, Mingsong
Machine Learning
Functional Analysis
This paper introduces a hypothesis space for deep learning based on deep neural networks (DNNs). By treating a DNN as a function of two variables - the input variable and the parameter variable - we consider the set of DNNs where the parameter variable belongs to a space of weight matrices and biases determined by a prescribed depth and layer widths. To construct a Banach space of functions of the input variable, we take the weak* closure of the linear span of this DNN set. We prove that the resulting Banach space is a reproducing kernel Banach space (RKBS) and explicitly construct its reproducing kernel. Furthermore, we investigate two learning models - regularized learning and the minimum norm interpolation (MNI) problem - within the RKBS framework by establishing representer theorems. These theorems reveal that the solutions to these learning problems can be expressed as a finite sum of kernel expansions based on training data.
title Hypothesis Spaces for Deep Learning
topic Machine Learning
Functional Analysis
url https://arxiv.org/abs/2403.03353