Bias-inducing geometries: an exactly solvable data model with fairness implications
Fuente:
arXiv
Salvato in:
| Autori principali: | Mannelli, Stefano Sarao, Gerace, Federica, Rostamzadeh, Negar, Saglietti, Luca |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2022
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
The Interplay of Data Structure and Imbalance in the Learning Dynamics of Diffusion Models
di: Nicoletti, Flavio, et al.
Pubblicazione: (2026)
di: Nicoletti, Flavio, et al.
Pubblicazione: (2026)
Tilting the Odds at the Lottery: the Interplay of Overparameterisation and Curricula in Neural Networks
di: Mannelli, Stefano Sarao, et al.
Pubblicazione: (2024)
di: Mannelli, Stefano Sarao, et al.
Pubblicazione: (2024)
Bias in Motion: Theoretical Insights into the Dynamics of Bias in SGD Training
di: Jain, Anchit, et al.
Pubblicazione: (2024)
di: Jain, Anchit, et al.
Pubblicazione: (2024)
Sharp description of local minima in the loss landscape of high-dimensional two-layer ReLU neural networks
di: Huang, Jie, et al.
Pubblicazione: (2026)
di: Huang, Jie, et al.
Pubblicazione: (2026)
Optimal Protocols for Continual Learning via Statistical Physics and Control Theory
di: Mori, Francesco, et al.
Pubblicazione: (2024)
di: Mori, Francesco, et al.
Pubblicazione: (2024)
The RL Perceptron: Generalisation Dynamics of Policy Learning in High Dimensions
di: Patel, Nishil, et al.
Pubblicazione: (2023)
di: Patel, Nishil, et al.
Pubblicazione: (2023)
An exactly solvable model for emergence and scaling laws in the multitask sparse parity problem
di: Nam, Yoonsoo, et al.
Pubblicazione: (2024)
di: Nam, Yoonsoo, et al.
Pubblicazione: (2024)
Why Do Animals Need Shaping? A Theory of Task Composition and Curriculum Learning
di: Lee, Jin Hwa, et al.
Pubblicazione: (2024)
di: Lee, Jin Hwa, et al.
Pubblicazione: (2024)
Isolating the hard core of phaseless inference: the Phase selection formulation
di: Straziota, Davide, et al.
Pubblicazione: (2025)
di: Straziota, Davide, et al.
Pubblicazione: (2025)
A solvable model of learning generative diffusion: theory and insights
di: Cui, Hugo, et al.
Pubblicazione: (2025)
di: Cui, Hugo, et al.
Pubblicazione: (2025)
Biased Generalization in Diffusion Models
di: Garnier-Brun, Jerome, et al.
Pubblicazione: (2026)
di: Garnier-Brun, Jerome, et al.
Pubblicazione: (2026)
Mapping of attention mechanisms to a generalized Potts model
di: Rende, Riccardo, et al.
Pubblicazione: (2023)
di: Rende, Riccardo, et al.
Pubblicazione: (2023)
Escape dynamics and implicit bias of one-pass SGD in overparameterized quadratic networks
di: Bocchi, Dario, et al.
Pubblicazione: (2026)
di: Bocchi, Dario, et al.
Pubblicazione: (2026)
How transformers learn structured data: insights from hierarchical filtering
di: Garnier-Brun, Jerome, et al.
Pubblicazione: (2024)
di: Garnier-Brun, Jerome, et al.
Pubblicazione: (2024)
The twin peaks of learning neural networks
di: Demyanenko, Elizaveta, et al.
Pubblicazione: (2024)
di: Demyanenko, Elizaveta, et al.
Pubblicazione: (2024)
Gaussian Universality of Perceptrons with Random Labels
di: Gerace, Federica, et al.
Pubblicazione: (2022)
di: Gerace, Federica, et al.
Pubblicazione: (2022)
A solvable high-dimensional model where nonlinear autoencoders learn structure invisible to PCA while test loss misaligns with generalization
di: Mendes, Vicente Conde, et al.
Pubblicazione: (2026)
di: Mendes, Vicente Conde, et al.
Pubblicazione: (2026)
Demystifying Spectral Bias on Real-World Data
di: Lavie, Itay, et al.
Pubblicazione: (2024)
di: Lavie, Itay, et al.
Pubblicazione: (2024)
How does training shape the Riemannian geometry of neural network representations?
di: Zavatone-Veth, Jacob A., et al.
Pubblicazione: (2023)
di: Zavatone-Veth, Jacob A., et al.
Pubblicazione: (2023)
Generalization performance of narrow one-hidden layer networks in the teacher-student setting
di: Ortiz, Rodrigo Pérez, et al.
Pubblicazione: (2025)
di: Ortiz, Rodrigo Pérez, et al.
Pubblicazione: (2025)
Towards Understanding Inductive Bias in Transformers: A View From Infinity
di: Lavie, Itay, et al.
Pubblicazione: (2024)
di: Lavie, Itay, et al.
Pubblicazione: (2024)
Algorithmic Task Capture, Computational Complexity, and Inductive Bias of Infinite Transformers
di: Davidovich, Orit, et al.
Pubblicazione: (2026)
di: Davidovich, Orit, et al.
Pubblicazione: (2026)
Initial Guessing Bias: How Untrained Networks Favor Some Classes
di: Francazi, Emanuele, et al.
Pubblicazione: (2023)
di: Francazi, Emanuele, et al.
Pubblicazione: (2023)
Properties of the geometry of solutions and capacity of multi-layer neural networks with Rectified Linear Units activations
di: Baldassi, Carlo, et al.
Pubblicazione: (2019)
di: Baldassi, Carlo, et al.
Pubblicazione: (2019)
Symmetry in language statistics shapes the geometry of model representations
di: Karkada, Dhruva, et al.
Pubblicazione: (2026)
di: Karkada, Dhruva, et al.
Pubblicazione: (2026)
Statistical mechanics of transfer learning in fully-connected networks in the proportional limit
di: Ingrosso, Alessandro, et al.
Pubblicazione: (2024)
di: Ingrosso, Alessandro, et al.
Pubblicazione: (2024)
Soft Quantization: Model Compression Via Weight Coupling
di: Bernstein, Daniel T., et al.
Pubblicazione: (2026)
di: Bernstein, Daniel T., et al.
Pubblicazione: (2026)
Emergence of Distortions in High-Dimensional Guided Diffusion Models
di: Ventura, Enrico, et al.
Pubblicazione: (2026)
di: Ventura, Enrico, et al.
Pubblicazione: (2026)
Deep Learning as Neural Low-Degree Filtering: A Spectral Theory of Hierarchical Feature Learning
di: Dandi, Yatin, et al.
Pubblicazione: (2026)
di: Dandi, Yatin, et al.
Pubblicazione: (2026)
Bilinear Sequence Regression: A Model for Learning from Long Sequences of High-dimensional Tokens
di: Erba, Vittorio, et al.
Pubblicazione: (2024)
di: Erba, Vittorio, et al.
Pubblicazione: (2024)
Asymptotics of SGD in Sequence-Single Index Models and Single-Layer Attention Networks
di: Arnaboldi, Luca, et al.
Pubblicazione: (2025)
di: Arnaboldi, Luca, et al.
Pubblicazione: (2025)
Autoregressive model path dependence near Ising criticality
di: Teoh, Yi Hong, et al.
Pubblicazione: (2024)
di: Teoh, Yi Hong, et al.
Pubblicazione: (2024)
Asymptotics of feature learning in two-layer networks after one gradient-step
di: Cui, Hugo, et al.
Pubblicazione: (2024)
di: Cui, Hugo, et al.
Pubblicazione: (2024)
Nadaraya-Watson kernel smoothing as a random energy model
di: Zavatone-Veth, Jacob A., et al.
Pubblicazione: (2024)
di: Zavatone-Veth, Jacob A., et al.
Pubblicazione: (2024)
Generative modeling through internal high-dimensional chaotic activity
di: Fournier, Samantha J., et al.
Pubblicazione: (2024)
di: Fournier, Samantha J., et al.
Pubblicazione: (2024)
Restoring balance: principled under/oversampling of data for optimal classification
di: Loffredo, Emanuele, et al.
Pubblicazione: (2024)
di: Loffredo, Emanuele, et al.
Pubblicazione: (2024)
Non-equilibrium active noise enhances generative memory in diffusion models
di: Behera, Agnish Kumar, et al.
Pubblicazione: (2024)
di: Behera, Agnish Kumar, et al.
Pubblicazione: (2024)
A Federated Many-to-One Hopfield model for associative Neural Networks
di: Alessandrelli, Andrea, et al.
Pubblicazione: (2026)
di: Alessandrelli, Andrea, et al.
Pubblicazione: (2026)
Optimal thresholds and algorithms for a model of multi-modal learning in high dimensions
di: Keup, Christian, et al.
Pubblicazione: (2024)
di: Keup, Christian, et al.
Pubblicazione: (2024)
Learning curves theory for hierarchically compositional data with power-law distributed features
di: Cagnetta, Francesco, et al.
Pubblicazione: (2025)
di: Cagnetta, Francesco, et al.
Pubblicazione: (2025)
Documenti analoghi
-
The Interplay of Data Structure and Imbalance in the Learning Dynamics of Diffusion Models
di: Nicoletti, Flavio, et al.
Pubblicazione: (2026) -
Tilting the Odds at the Lottery: the Interplay of Overparameterisation and Curricula in Neural Networks
di: Mannelli, Stefano Sarao, et al.
Pubblicazione: (2024) -
Bias in Motion: Theoretical Insights into the Dynamics of Bias in SGD Training
di: Jain, Anchit, et al.
Pubblicazione: (2024) -
Sharp description of local minima in the loss landscape of high-dimensional two-layer ReLU neural networks
di: Huang, Jie, et al.
Pubblicazione: (2026) -
Optimal Protocols for Continual Learning via Statistical Physics and Control Theory
di: Mori, Francesco, et al.
Pubblicazione: (2024)