Inheritance Between Feedforward and Convolutional Networks via Model Projection

Fuente: arXiv
Gespeichert in:
Bibliographische Detailangaben
Hauptverfasser: Ewen, Nicolas, Diaz-Rodriguez, Jairo, Ramsay, Kelly
Format: Preprint
Veröffentlicht: 2026
Schlagworte:
Online-Zugang:
Tags: Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
_version_ 1866912882762973184
author Ewen, Nicolas
Diaz-Rodriguez, Jairo
Ramsay, Kelly
author_facet Ewen, Nicolas
Diaz-Rodriguez, Jairo
Ramsay, Kelly
contents Techniques for feedforward networks (FFNs) and convolutional networks (CNNs) are frequently reused across families, but the relationship between the underlying model classes is rarely made explicit. We introduce a unified node-level formalization with tensor-valued activations and show that generalized feedforward networks form a strict subset of generalized convolutional networks. Motivated by the mismatch in per-input parameterization between the two families, we propose model projection, a parameter-efficient transfer learning method for CNNs that freezes pretrained per-input-channel filters and learns a single scalar gate for each (output channel, input channel) contribution. Projection keeps all convolutional layers adaptable to downstream tasks while substantially reducing the number of trained parameters in convolutional layers. We prove that projected nodes take the generalized FFN form, enabling projected CNNs to inherit feedforward techniques that do not rely on homogeneous layer inputs. Experiments across multiple ImageNet-pretrained backbones and several downstream image classification datasets show that model projection is a strong transfer learning baseline under simple training recipes.
format Preprint
id arxiv_https___arxiv_org_abs_2602_06245
institution arXiv
publishDate 2026
record_format arxiv
spellingShingle Inheritance Between Feedforward and Convolutional Networks via Model Projection
Ewen, Nicolas
Diaz-Rodriguez, Jairo
Ramsay, Kelly
Machine Learning
Techniques for feedforward networks (FFNs) and convolutional networks (CNNs) are frequently reused across families, but the relationship between the underlying model classes is rarely made explicit. We introduce a unified node-level formalization with tensor-valued activations and show that generalized feedforward networks form a strict subset of generalized convolutional networks. Motivated by the mismatch in per-input parameterization between the two families, we propose model projection, a parameter-efficient transfer learning method for CNNs that freezes pretrained per-input-channel filters and learns a single scalar gate for each (output channel, input channel) contribution. Projection keeps all convolutional layers adaptable to downstream tasks while substantially reducing the number of trained parameters in convolutional layers. We prove that projected nodes take the generalized FFN form, enabling projected CNNs to inherit feedforward techniques that do not rely on homogeneous layer inputs. Experiments across multiple ImageNet-pretrained backbones and several downstream image classification datasets show that model projection is a strong transfer learning baseline under simple training recipes.
title Inheritance Between Feedforward and Convolutional Networks via Model Projection
topic Machine Learning
url https://arxiv.org/abs/2602.06245