On Monotonicity in AI Alignment
Fuente:
arXiv
Saved in:
| Main Authors: | Bareilles, Gilles, Fageot, Julien, Hoang, Lê-Nguyên, Blanchard, Peva, Bouaziz, Wassim, Rouault, Sébastien, El-Mhamdi, El-Mahdi |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Byzantine Machine Learning: MultiKrum and an optimal notion of robustness
by: Bareilles, Gilles, et al.
Published: (2026)
by: Bareilles, Gilles, et al.
Published: (2026)
Generalizing while preserving monotonicity in comparison-based preference learning models
by: Fageot, Julien, et al.
Published: (2025)
by: Fageot, Julien, et al.
Published: (2025)
On Goodhart's law, with an application to value alignment
by: El-Mhamdi, El-Mahdi, et al.
Published: (2024)
by: El-Mhamdi, El-Mahdi, et al.
Published: (2024)
The Strong, Weak and Benign Goodhart's law. An independence-free and paradigm-agnostic formalisation
by: Majka, Adrien, et al.
Published: (2025)
by: Majka, Adrien, et al.
Published: (2025)
Targeted Data Poisoning for Black-Box Audio Datasets Ownership Verification
by: Bouaziz, Wassim, et al.
Published: (2025)
by: Bouaziz, Wassim, et al.
Published: (2025)
Data Taggants: Dataset Ownership Verification via Harmless Targeted Data Poisoning
by: Bouaziz, Wassim, et al.
Published: (2024)
by: Bouaziz, Wassim, et al.
Published: (2024)
Inverting Gradient Attacks Makes Powerful Data Poisoning
by: Bouaziz, Wassim, et al.
Published: (2024)
by: Bouaziz, Wassim, et al.
Published: (2024)
Winter Soldier: Backdooring Language Models at Pre-Training with Indirect Data Poisoning
by: Bouaziz, Wassim, et al.
Published: (2025)
by: Bouaziz, Wassim, et al.
Published: (2025)
Piecewise Polynomial Regression of Tame Functions via Integer Programming
by: Bareilles, Gilles, et al.
Published: (2023)
by: Bareilles, Gilles, et al.
Published: (2023)
Generalized Bradley-Terry Models for Score Estimation from Paired Comparisons
by: Fageot, Julien, et al.
Published: (2023)
by: Fageot, Julien, et al.
Published: (2023)
Convergence of Statistical Estimators via Mutual Information Bounds
by: Khribch, El Mahdi, et al.
Published: (2024)
by: Khribch, El Mahdi, et al.
Published: (2024)
The loss landscape of deep linear neural networks: a second-order analysis
by: Achour, El Mehdi, et al.
Published: (2021)
by: Achour, El Mehdi, et al.
Published: (2021)
Statistical learning on measures: an application to persistence diagrams
by: Hacquard, Olympio, et al.
Published: (2023)
by: Hacquard, Olympio, et al.
Published: (2023)
Pseudo-Observations and Super Learner for the Estimation of the Restricted Mean Survival Time
by: Cwiling, Ariane, et al.
Published: (2024)
by: Cwiling, Ariane, et al.
Published: (2024)
Meta-Learning and representation learner: A short theoretical note
by: Bouchattaoui, Mouad El
Published: (2024)
by: Bouchattaoui, Mouad El
Published: (2024)
Distributionally-Constrained Adversaries in Online Learning
by: Blanchard, Moïse, et al.
Published: (2025)
by: Blanchard, Moïse, et al.
Published: (2025)
On Learning-Curve Monotonicity for Maximum Likelihood Estimators
by: Sellke, Mark, et al.
Published: (2025)
by: Sellke, Mark, et al.
Published: (2025)
Optimal Excess Risk Bounds for Empirical Risk Minimization on $p$-Norm Linear Regression
by: Hanchi, Ayoub El, et al.
Published: (2023)
by: Hanchi, Ayoub El, et al.
Published: (2023)
Approximation of Maximally Monotone Operators : A Graph Convergence Perspective
by: Furuya, Takashi, et al.
Published: (2026)
by: Furuya, Takashi, et al.
Published: (2026)
Approaching the Harm of Gradient Attacks While Only Flipping Labels
by: El-Kabid, Abdessamad, et al.
Published: (2025)
by: El-Kabid, Abdessamad, et al.
Published: (2025)
A Geometric Analysis of PCA
by: Hanchi, Ayoub El, et al.
Published: (2025)
by: Hanchi, Ayoub El, et al.
Published: (2025)
Conformal Risk Control for Non-Monotonic Losses
by: Angelopoulos, Anastasios N.
Published: (2026)
by: Angelopoulos, Anastasios N.
Published: (2026)
General reproducing properties in RKHS with application to derivative and integral operators
by: El-Boukkouri, Fatima-Zahrae, et al.
Published: (2025)
by: El-Boukkouri, Fatima-Zahrae, et al.
Published: (2025)
Variational Representations of Annealing Paths: Bregman Information under Monotonic Embedding
by: Brekelmans, Rob, et al.
Published: (2022)
by: Brekelmans, Rob, et al.
Published: (2022)
Minimax Linear Regression under the Quantile Risk
by: Hanchi, Ayoub El, et al.
Published: (2024)
by: Hanchi, Ayoub El, et al.
Published: (2024)
On the Efficiency of ERM in Feature Learning
by: Hanchi, Ayoub El, et al.
Published: (2024)
by: Hanchi, Ayoub El, et al.
Published: (2024)
Consistency and Inconsistency in $K$-Means Clustering
by: Blanchard, Moïse, et al.
Published: (2025)
by: Blanchard, Moïse, et al.
Published: (2025)
Robust Bayesian Inference via Variational Approximations of Generalized Rho-Posteriors
by: Khribch, EL Mahdi, et al.
Published: (2026)
by: Khribch, EL Mahdi, et al.
Published: (2026)
Monotone Curve Estimation via Convex Duality
by: Lim, Tongseok, et al.
Published: (2025)
by: Lim, Tongseok, et al.
Published: (2025)
High-probability zeroth-order online convex optimisation beyond Euclidean geometry
by: Janz, David, et al.
Published: (2025)
by: Janz, David, et al.
Published: (2025)
Optimal community detection in dense bipartite graphs
by: Chhor, Julien, et al.
Published: (2025)
by: Chhor, Julien, et al.
Published: (2025)
Convex SGD: Generalization Without Early Stopping
by: Hendrickx, Julien, et al.
Published: (2024)
by: Hendrickx, Julien, et al.
Published: (2024)
Improving Minimax Estimation Rates for Contaminated Mixture of Multinomial Logistic Experts via Expert Heterogeneity
by: Yan, Fanqi, et al.
Published: (2026)
by: Yan, Fanqi, et al.
Published: (2026)
Maximizing the Potential of Synthetic Data: Insights from Random Matrix Theory
by: Firdoussi, Aymane El, et al.
Published: (2024)
by: Firdoussi, Aymane El, et al.
Published: (2024)
Gaussian-Smoothed Sliced Probability Divergences
by: Alaya, Mokhtar Z., et al.
Published: (2024)
by: Alaya, Mokhtar Z., et al.
Published: (2024)
Minimax optimal submatrix detection: Sharp non-asymptotic rates
by: Knight, Parker, et al.
Published: (2026)
by: Knight, Parker, et al.
Published: (2026)
A Unified Framework for Variable Selection in Model-Based Clustering with Missing Not at Random
by: Ho, Binh H., et al.
Published: (2025)
by: Ho, Binh H., et al.
Published: (2025)
Scalable and adaptive prediction bands with kernel sum-of-squares
by: Allain, Louis, et al.
Published: (2025)
by: Allain, Louis, et al.
Published: (2025)
Robustly Learning Monotone Generalized Linear Models via Data Augmentation
by: Zarifis, Nikos, et al.
Published: (2025)
by: Zarifis, Nikos, et al.
Published: (2025)
Robust Alignment via Partial Gromov-Wasserstein Distances
by: Gong, Xiaoyun, et al.
Published: (2025)
by: Gong, Xiaoyun, et al.
Published: (2025)
Similar Items
-
Byzantine Machine Learning: MultiKrum and an optimal notion of robustness
by: Bareilles, Gilles, et al.
Published: (2026) -
Generalizing while preserving monotonicity in comparison-based preference learning models
by: Fageot, Julien, et al.
Published: (2025) -
On Goodhart's law, with an application to value alignment
by: El-Mhamdi, El-Mahdi, et al.
Published: (2024) -
The Strong, Weak and Benign Goodhart's law. An independence-free and paradigm-agnostic formalisation
by: Majka, Adrien, et al.
Published: (2025) -
Targeted Data Poisoning for Black-Box Audio Datasets Ownership Verification
by: Bouaziz, Wassim, et al.
Published: (2025)