The Strong, Weak and Benign Goodhart's law. An independence-free and paradigm-agnostic formalisation
Fuente:
arXiv
Salvato in:
| Autori principali: | Majka, Adrien, El-Mhamdi, El-Mahdi |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
On Goodhart's law, with an application to value alignment
di: El-Mhamdi, El-Mahdi, et al.
Pubblicazione: (2024)
di: El-Mhamdi, El-Mahdi, et al.
Pubblicazione: (2024)
Byzantine Machine Learning: MultiKrum and an optimal notion of robustness
di: Bareilles, Gilles, et al.
Pubblicazione: (2026)
di: Bareilles, Gilles, et al.
Pubblicazione: (2026)
On Monotonicity in AI Alignment
di: Bareilles, Gilles, et al.
Pubblicazione: (2025)
di: Bareilles, Gilles, et al.
Pubblicazione: (2025)
Convergence of Statistical Estimators via Mutual Information Bounds
di: Khribch, El Mahdi, et al.
Pubblicazione: (2024)
di: Khribch, El Mahdi, et al.
Pubblicazione: (2024)
Beyond Benign Overfitting in Nadaraya-Watson Interpolators
di: Barzilai, Daniel, et al.
Pubblicazione: (2025)
di: Barzilai, Daniel, et al.
Pubblicazione: (2025)
Meta-Learning and representation learner: A short theoretical note
di: Bouchattaoui, Mouad El
Pubblicazione: (2024)
di: Bouchattaoui, Mouad El
Pubblicazione: (2024)
Transfer Learning for Benign Overfitting in High-Dimensional Linear Regression
di: Kim, Yeichan, et al.
Pubblicazione: (2025)
di: Kim, Yeichan, et al.
Pubblicazione: (2025)
Learning Curves and Benign Overfitting of Spectral Algorithms in Large Dimensions
di: Lu, Weihao, et al.
Pubblicazione: (2026)
di: Lu, Weihao, et al.
Pubblicazione: (2026)
Benign Overfitting in Time Series Linear Models with Over-Parameterization
di: Nakakita, Shogo, et al.
Pubblicazione: (2022)
di: Nakakita, Shogo, et al.
Pubblicazione: (2022)
Mind the spikes: Benign overfitting of kernels and neural networks in fixed dimension
di: Haas, Moritz, et al.
Pubblicazione: (2023)
di: Haas, Moritz, et al.
Pubblicazione: (2023)
Dimension-agnostic inference using cross U-statistics
di: Kim, Ilmun, et al.
Pubblicazione: (2020)
di: Kim, Ilmun, et al.
Pubblicazione: (2020)
Benign Overfitting under Learning Rate Conditions for $α$ Sub-exponential Input
di: Okudo, Kota, et al.
Pubblicazione: (2024)
di: Okudo, Kota, et al.
Pubblicazione: (2024)
A general framework for inference on algorithm-agnostic variable importance
di: Williamson, Brian D., et al.
Pubblicazione: (2020)
di: Williamson, Brian D., et al.
Pubblicazione: (2020)
Model-agnostic information transfer and fusion for classification with label noise
di: Guojun, Zhu, et al.
Pubblicazione: (2026)
di: Guojun, Zhu, et al.
Pubblicazione: (2026)
Optimal Excess Risk Bounds for Empirical Risk Minimization on $p$-Norm Linear Regression
di: Hanchi, Ayoub El, et al.
Pubblicazione: (2023)
di: Hanchi, Ayoub El, et al.
Pubblicazione: (2023)
Approaching the Harm of Gradient Attacks While Only Flipping Labels
di: El-Kabid, Abdessamad, et al.
Pubblicazione: (2025)
di: El-Kabid, Abdessamad, et al.
Pubblicazione: (2025)
A Geometric Analysis of PCA
di: Hanchi, Ayoub El, et al.
Pubblicazione: (2025)
di: Hanchi, Ayoub El, et al.
Pubblicazione: (2025)
The loss landscape of deep linear neural networks: a second-order analysis
di: Achour, El Mehdi, et al.
Pubblicazione: (2021)
di: Achour, El Mehdi, et al.
Pubblicazione: (2021)
Benign Overfitting without Linearity: Neural Network Classifiers Trained by Gradient Descent for Noisy Linear Data
di: Frei, Spencer, et al.
Pubblicazione: (2022)
di: Frei, Spencer, et al.
Pubblicazione: (2022)
It's Hard to Be Normal: The Impact of Noise on Structure-agnostic Estimation
di: Jin, Jikai, et al.
Pubblicazione: (2025)
di: Jin, Jikai, et al.
Pubblicazione: (2025)
General reproducing properties in RKHS with application to derivative and integral operators
di: El-Boukkouri, Fatima-Zahrae, et al.
Pubblicazione: (2025)
di: El-Boukkouri, Fatima-Zahrae, et al.
Pubblicazione: (2025)
Optimal structure learning and conditional independence testing
di: Gao, Ming, et al.
Pubblicazione: (2025)
di: Gao, Ming, et al.
Pubblicazione: (2025)
Minimax Linear Regression under the Quantile Risk
di: Hanchi, Ayoub El, et al.
Pubblicazione: (2024)
di: Hanchi, Ayoub El, et al.
Pubblicazione: (2024)
On the Efficiency of ERM in Feature Learning
di: Hanchi, Ayoub El, et al.
Pubblicazione: (2024)
di: Hanchi, Ayoub El, et al.
Pubblicazione: (2024)
Characterizing the Generalization Error of Random Feature Regression with Arbitrary Data-Augmentation
di: Morisset, Lucas, et al.
Pubblicazione: (2026)
di: Morisset, Lucas, et al.
Pubblicazione: (2026)
Structure-agnostic Optimality of Doubly Robust Learning for Treatment Effect Estimation
di: Jin, Jikai, et al.
Pubblicazione: (2024)
di: Jin, Jikai, et al.
Pubblicazione: (2024)
Universality of Benign Overfitting in Binary Linear Classification
di: Hashimoto, Ichiro, et al.
Pubblicazione: (2025)
di: Hashimoto, Ichiro, et al.
Pubblicazione: (2025)
Asymptotically Optimal Problem-Dependent Bandit Policies for Transfer Learning
di: Prevost, Adrien, et al.
Pubblicazione: (2025)
di: Prevost, Adrien, et al.
Pubblicazione: (2025)
Higher-Order Regularization Learning on Hypergraphs
di: Weihs, Adrien, et al.
Pubblicazione: (2025)
di: Weihs, Adrien, et al.
Pubblicazione: (2025)
Analysis of Semi-Supervised Learning on Hypergraphs
di: Weihs, Adrien, et al.
Pubblicazione: (2025)
di: Weihs, Adrien, et al.
Pubblicazione: (2025)
Robust Bayesian Inference via Variational Approximations of Generalized Rho-Posteriors
di: Khribch, EL Mahdi, et al.
Pubblicazione: (2026)
di: Khribch, EL Mahdi, et al.
Pubblicazione: (2026)
Universality laws for Gaussian mixtures in generalized linear models
di: Dandi, Yatin, et al.
Pubblicazione: (2023)
di: Dandi, Yatin, et al.
Pubblicazione: (2023)
High-probability zeroth-order online convex optimisation beyond Euclidean geometry
di: Janz, David, et al.
Pubblicazione: (2025)
di: Janz, David, et al.
Pubblicazione: (2025)
Strong identifiability and parameter learning in regression with heterogeneous response
di: Do, Dat, et al.
Pubblicazione: (2022)
di: Do, Dat, et al.
Pubblicazione: (2022)
Maximizing the Potential of Synthetic Data: Insights from Random Matrix Theory
di: Firdoussi, Aymane El, et al.
Pubblicazione: (2024)
di: Firdoussi, Aymane El, et al.
Pubblicazione: (2024)
Benign overfitting in Fixed Dimension via Physics-Informed Learning with Smooth Inductive Bias
di: Wong, Honam, et al.
Pubblicazione: (2024)
di: Wong, Honam, et al.
Pubblicazione: (2024)
Non-Asymptotic Analysis of Data Augmentation for Precision Matrix Estimation
di: Morisset, Lucas, et al.
Pubblicazione: (2025)
di: Morisset, Lucas, et al.
Pubblicazione: (2025)
Sampling from multi-modal distributions with polynomial query complexity in fixed dimension via reverse diffusion
di: Vacher, Adrien, et al.
Pubblicazione: (2024)
di: Vacher, Adrien, et al.
Pubblicazione: (2024)
Generalization Error Curves for Analytic Spectral Algorithms under Power-law Decay
di: Li, Yicheng, et al.
Pubblicazione: (2024)
di: Li, Yicheng, et al.
Pubblicazione: (2024)
Inferring the finest pattern of mutual independence from data
di: Marrelec, G., et al.
Pubblicazione: (2023)
di: Marrelec, G., et al.
Pubblicazione: (2023)
Documenti analoghi
-
On Goodhart's law, with an application to value alignment
di: El-Mhamdi, El-Mahdi, et al.
Pubblicazione: (2024) -
Byzantine Machine Learning: MultiKrum and an optimal notion of robustness
di: Bareilles, Gilles, et al.
Pubblicazione: (2026) -
On Monotonicity in AI Alignment
di: Bareilles, Gilles, et al.
Pubblicazione: (2025) -
Convergence of Statistical Estimators via Mutual Information Bounds
di: Khribch, El Mahdi, et al.
Pubblicazione: (2024) -
Beyond Benign Overfitting in Nadaraya-Watson Interpolators
di: Barzilai, Daniel, et al.
Pubblicazione: (2025)