Relax and penalize: a new bilevel approach to mixed-binary hyperparameter optimization
Fuente:
arXiv
Saved in:
| Main Authors: | Venturini, Sara, de Santis, Marianna, Patracone, Jordan, Rinaldi, Francesco, Salzo, Saverio, Schmidt, Martin |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Iteration Complexity of Frank-Wolfe and Its Variants for Bilevel Optimization
by: Palmieri, Anthony, et al.
Published: (2026)
by: Palmieri, Anthony, et al.
Published: (2026)
An adaptively inexact first-order method for bilevel optimization with application to hyperparameter learning
by: Salehi, Mohammad Sadegh, et al.
Published: (2023)
by: Salehi, Mohammad Sadegh, et al.
Published: (2023)
Convergence Properties of Stochastic Hypergradients
by: Grazzi, Riccardo, et al.
Published: (2020)
by: Grazzi, Riccardo, et al.
Published: (2020)
AdaGrad-Diff: A New Version of the Adaptive Gradient Algorithm
by: Bojovic, Matia, et al.
Published: (2026)
by: Bojovic, Matia, et al.
Published: (2026)
Nonsmooth Implicit Differentiation: Deterministic and Stochastic Convergence Rates
by: Grazzi, Riccardo, et al.
Published: (2024)
by: Grazzi, Riccardo, et al.
Published: (2024)
Conformal Online Learning of Deep Koopman Linear Embeddings
by: Gao, Ben, et al.
Published: (2025)
by: Gao, Ben, et al.
Published: (2025)
Variance reduction techniques for stochastic proximal point algorithms
by: Traoré, Cheik, et al.
Published: (2023)
by: Traoré, Cheik, et al.
Published: (2023)
Be aware of overfitting by hyperparameter optimization!
by: Tetko, Igor V., et al.
Published: (2024)
by: Tetko, Igor V., et al.
Published: (2024)
The iterates of FISTA convergence even under inexact computations and stochastic gradients
by: Salzo, Saverio
Published: (2025)
by: Salzo, Saverio
Published: (2025)
High Probability Bounds for Stochastic Subgradient Schemes with Heavy Tailed Noise
by: Parletta, Daniela A., et al.
Published: (2022)
by: Parletta, Daniela A., et al.
Published: (2022)
A Hessian-informed hyperparameter optimization for differential learning rate
by: Xu, Shiyun, et al.
Published: (2025)
by: Xu, Shiyun, et al.
Published: (2025)
Two-step hyperparameter optimization method: Accelerating hyperparameter search by using a fraction of a training dataset
by: Yu, Sungduk, et al.
Published: (2023)
by: Yu, Sungduk, et al.
Published: (2023)
Fuzzy hyperparameters update in a second order optimization
by: Bensadok, Abdelaziz, et al.
Published: (2024)
by: Bensadok, Abdelaziz, et al.
Published: (2024)
Beyond algorithm hyperparameters: on preprocessing hyperparameters and associated pitfalls in machine learning applications
by: Sauer, Christina, et al.
Published: (2024)
by: Sauer, Christina, et al.
Published: (2024)
Bayesian autoregression to optimize temporal Matérn kernel Gaussian process hyperparameters
by: Kouw, Wouter M.
Published: (2025)
by: Kouw, Wouter M.
Published: (2025)
A two-step sequential approach for hyperparameter selection in finite context models
by: Contente, José, et al.
Published: (2026)
by: Contente, José, et al.
Published: (2026)
Towards hyperparameter-free optimization with differential privacy
by: Bu, Zhiqi, et al.
Published: (2025)
by: Bu, Zhiqi, et al.
Published: (2025)
Application-oriented automatic hyperparameter optimization for spiking neural network prototyping
by: Fra, Vittorio
Published: (2025)
by: Fra, Vittorio
Published: (2025)
A framework for bilevel optimization that enables stochastic and global variance reduction algorithms
by: Dagréou, Mathieu, et al.
Published: (2022)
by: Dagréou, Mathieu, et al.
Published: (2022)
An algorithmic framework for the optimization of deep neural networks architectures and hyperparameters
by: Keisler, Julie, et al.
Published: (2023)
by: Keisler, Julie, et al.
Published: (2023)
Regularized boosting with an increasing coefficient magnitude stop criterion as meta-learner in hyperparameter optimization stacking ensemble
by: Fdez-Díaz, Laura, et al.
Published: (2024)
by: Fdez-Díaz, Laura, et al.
Published: (2024)
An experimental comparative study of backpropagation and alternatives for training binary neural networks for image classification
by: Crulis, Ben, et al.
Published: (2024)
by: Crulis, Ben, et al.
Published: (2024)
LancBiO: dynamic Lanczos-aided bilevel optimization via Krylov subspace
by: Yang, Yan, et al.
Published: (2024)
by: Yang, Yan, et al.
Published: (2024)
Bilevel optimization for learning hyperparameters: Application to solving PDEs and inverse problems with Gaussian processes
by: Nelsen, Nicholas H., et al.
Published: (2025)
by: Nelsen, Nicholas H., et al.
Published: (2025)
Decision-focused predictions via pessimistic bilevel optimization: complexity and algorithms
by: Bucarey, Víctor, et al.
Published: (2023)
by: Bucarey, Víctor, et al.
Published: (2023)
Deep learning from strongly mixing observations: Sparse-penalized regularization and minimax optimality
by: Kengne, William, et al.
Published: (2024)
by: Kengne, William, et al.
Published: (2024)
Effect of hyperparameters on variable selection in random forests
by: Fouodo, Cesaire J. K., et al.
Published: (2023)
by: Fouodo, Cesaire J. K., et al.
Published: (2023)
Offline-to-online hyperparameter transfer for stochastic bandits
by: Sharma, Dravyansh, et al.
Published: (2025)
by: Sharma, Dravyansh, et al.
Published: (2025)
Calibrating dimension reduction hyperparameters in the presence of noise
by: Lin, Justin, et al.
Published: (2023)
by: Lin, Justin, et al.
Published: (2023)
Selecting time-series hyperparameters with the artificial jackknife
by: Pellegrino, Filippo
Published: (2020)
by: Pellegrino, Filippo
Published: (2020)
Classification under strategic adversary manipulation using pessimistic bilevel optimisation
by: Benfield, David, et al.
Published: (2024)
by: Benfield, David, et al.
Published: (2024)
Solving bilevel optimization via sequential minimax optimization
by: Lu, Zhaosong, et al.
Published: (2025)
by: Lu, Zhaosong, et al.
Published: (2025)
How far away are truly hyperparameter-free learning algorithms?
by: Kasimbeg, Priya, et al.
Published: (2025)
by: Kasimbeg, Priya, et al.
Published: (2025)
Relaxed Gaussian process interpolation: a goal-oriented approach to Bayesian optimization
by: Petit, Sébastien, et al.
Published: (2022)
by: Petit, Sébastien, et al.
Published: (2022)
Relaxation methods for pessimistic bilevel optimization
by: Benchouk, Imane, et al.
Published: (2024)
by: Benchouk, Imane, et al.
Published: (2024)
Better Trees: An empirical study on hyperparameter tuning of classification decision tree induction algorithms
by: Mantovani, Rafael Gomes, et al.
Published: (2018)
by: Mantovani, Rafael Gomes, et al.
Published: (2018)
First-order penalty methods for bilevel optimization
by: Lu, Zhaosong, et al.
Published: (2023)
by: Lu, Zhaosong, et al.
Published: (2023)
Auto Researching, not hyperparameter tuning: Convergence Analysis of 10,000 Experiments
by: Li, Xiaoyi
Published: (2026)
by: Li, Xiaoyi
Published: (2026)
$μ$pscaling small models: Principled warm starts and hyperparameter transfer
by: Ma, Yuxin, et al.
Published: (2026)
by: Ma, Yuxin, et al.
Published: (2026)
Time-optimal neural feedback control of nilpotent systems as a binary classification problem
by: Bicego, Sara, et al.
Published: (2025)
by: Bicego, Sara, et al.
Published: (2025)
Similar Items
-
Iteration Complexity of Frank-Wolfe and Its Variants for Bilevel Optimization
by: Palmieri, Anthony, et al.
Published: (2026) -
An adaptively inexact first-order method for bilevel optimization with application to hyperparameter learning
by: Salehi, Mohammad Sadegh, et al.
Published: (2023) -
Convergence Properties of Stochastic Hypergradients
by: Grazzi, Riccardo, et al.
Published: (2020) -
AdaGrad-Diff: A New Version of the Adaptive Gradient Algorithm
by: Bojovic, Matia, et al.
Published: (2026) -
Nonsmooth Implicit Differentiation: Deterministic and Stochastic Convergence Rates
by: Grazzi, Riccardo, et al.
Published: (2024)