Tree-Preconditioned Differentiable Optimization and Axioms as Layers
Fuente:
arXiv
Guardado en:
| Autor principal: | Liao, Yuexin |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Differentiable Distributionally Robust Optimization Layers
por: Ma, Xutao, et al.
Publicado: (2024)
por: Ma, Xutao, et al.
Publicado: (2024)
Preconditioning Benefits of Spectral Orthogonalization in Muon
por: Ma, Jianhao, et al.
Publicado: (2026)
por: Ma, Jianhao, et al.
Publicado: (2026)
Uncovering Symmetry Transfer in Large Language Models via Layer-Peeled Optimization
por: Du, Zhehang, et al.
Publicado: (2026)
por: Du, Zhehang, et al.
Publicado: (2026)
PDE Control Gym: A Benchmark for Data-Driven Boundary Control of Partial Differential Equations
por: Bhan, Luke, et al.
Publicado: (2024)
por: Bhan, Luke, et al.
Publicado: (2024)
Differentiable Optimization for Deep Learning-Enhanced DC Approximation of AC Optimal Power Flow
por: Rosemberg, Andrew, et al.
Publicado: (2025)
por: Rosemberg, Andrew, et al.
Publicado: (2025)
Transformers Can Implement Preconditioned Richardson Iteration for In-Context Gaussian Kernel Regression
por: Yan, Mingsong, et al.
Publicado: (2026)
por: Yan, Mingsong, et al.
Publicado: (2026)
Stability of Transformers under Layer Normalization
por: Kan, Kelvin, et al.
Publicado: (2025)
por: Kan, Kelvin, et al.
Publicado: (2025)
Differentiable Nonlinear Model Predictive Control
por: Frey, Jonathan, et al.
Publicado: (2025)
por: Frey, Jonathan, et al.
Publicado: (2025)
DT-PBO: an Interpretable Tree-based Surrogate Model for Preferential Bayesian Optimization
por: Leenders, Nick, et al.
Publicado: (2025)
por: Leenders, Nick, et al.
Publicado: (2025)
Implicit Regularization of Gradient Flow on One-Layer Softmax Attention
por: Sheen, Heejune, et al.
Publicado: (2024)
por: Sheen, Heejune, et al.
Publicado: (2024)
High-order expansion of Neural Ordinary Differential Equations flows
por: Izzo, Dario, et al.
Publicado: (2025)
por: Izzo, Dario, et al.
Publicado: (2025)
One-Layer Transformer Provably Learns One-Nearest Neighbor In Context
por: Li, Zihao, et al.
Publicado: (2024)
por: Li, Zihao, et al.
Publicado: (2024)
SGD at the Edge of Stability: The Stochastic Sharpness Gap
por: Liao, Fangshuo, et al.
Publicado: (2026)
por: Liao, Fangshuo, et al.
Publicado: (2026)
One Rank at a Time: Cascading Error Dynamics in Sequential Learning
por: Vandchali, Mahtab Alizadeh, et al.
Publicado: (2025)
por: Vandchali, Mahtab Alizadeh, et al.
Publicado: (2025)
Diagonalisation SGD: Fast & Convergent SGD for Non-Differentiable Models via Reparameterisation and Smoothing
por: Wagner, Dominik, et al.
Publicado: (2024)
por: Wagner, Dominik, et al.
Publicado: (2024)
Optimizing the Optimizer for Physics-Informed Neural Networks and Kolmogorov-Arnold Networks
por: Kiyani, Elham, et al.
Publicado: (2025)
por: Kiyani, Elham, et al.
Publicado: (2025)
The Newton-Muon Optimizer
por: Du, Zhehang, et al.
Publicado: (2026)
por: Du, Zhehang, et al.
Publicado: (2026)
Riemannian Bilevel Optimization
por: Dutta, Sanchayan, et al.
Publicado: (2024)
por: Dutta, Sanchayan, et al.
Publicado: (2024)
Optimizer-Model Consistency: Full Finetuning with the Same Optimizer as Pretraining Forgets Less
por: Liu, Yuxing, et al.
Publicado: (2026)
por: Liu, Yuxing, et al.
Publicado: (2026)
MetaOptimize: A Framework for Optimizing Step Sizes and Other Meta-parameters
por: Sharifnassab, Arsalan, et al.
Publicado: (2024)
por: Sharifnassab, Arsalan, et al.
Publicado: (2024)
From Large Language Models and Optimization to Decision Optimization CoPilot: A Research Manifesto
por: Wasserkrug, Segev, et al.
Publicado: (2024)
por: Wasserkrug, Segev, et al.
Publicado: (2024)
Global Convergence and Rich Feature Learning in $L$-Layer Infinite-Width Neural Networks under $μ$P Parametrization
por: Chen, Zixiang, et al.
Publicado: (2025)
por: Chen, Zixiang, et al.
Publicado: (2025)
Client-Centric Federated Adaptive Optimization
por: Sun, Jianhui, et al.
Publicado: (2025)
por: Sun, Jianhui, et al.
Publicado: (2025)
On the Condition Number Dependency in Bilevel Optimization
por: Chen, Lesi, et al.
Publicado: (2025)
por: Chen, Lesi, et al.
Publicado: (2025)
Budget-aware Auto Optimizer Configurator
por: Liu, Kang, et al.
Publicado: (2026)
por: Liu, Kang, et al.
Publicado: (2026)
Jacobian Descent for Multi-Objective Optimization
por: Quinton, Pierre, et al.
Publicado: (2024)
por: Quinton, Pierre, et al.
Publicado: (2024)
Optimization and Generalization Guarantees for Weight Normalization
por: Cisneros-Velarde, Pedro, et al.
Publicado: (2024)
por: Cisneros-Velarde, Pedro, et al.
Publicado: (2024)
Neural Solver Selection for Combinatorial Optimization
por: Gao, Chengrui, et al.
Publicado: (2024)
por: Gao, Chengrui, et al.
Publicado: (2024)
Dynamic Memory Based Adaptive Optimization
por: Szegedy, Balázs, et al.
Publicado: (2024)
por: Szegedy, Balázs, et al.
Publicado: (2024)
Zeroth-Order Optimization Finds Flat Minima
por: Zhang, Liang, et al.
Publicado: (2025)
por: Zhang, Liang, et al.
Publicado: (2025)
Search-Optimized Quantization in Biomedical Ontology Alignment
por: Bouaggad, Oussama, et al.
Publicado: (2025)
por: Bouaggad, Oussama, et al.
Publicado: (2025)
A Minimalist Bayesian Framework for Stochastic Optimization
por: Wang, Kaizheng
Publicado: (2025)
por: Wang, Kaizheng
Publicado: (2025)
Understanding Optimization in Deep Learning with Central Flows
por: Cohen, Jeremy M., et al.
Publicado: (2024)
por: Cohen, Jeremy M., et al.
Publicado: (2024)
Hierarchical Mixture-of-Experts with Two-Stage Optimization
por: Molodtsov, Gleb, et al.
Publicado: (2026)
por: Molodtsov, Gleb, et al.
Publicado: (2026)
Bayesian Optimization for Hyperparameters Tuning in Neural Networks
por: Onorato, Gabriele
Publicado: (2024)
por: Onorato, Gabriele
Publicado: (2024)
Intersectional Fairness via Mixed-Integer Optimization
por: Němeček, Jiří, et al.
Publicado: (2026)
por: Němeček, Jiří, et al.
Publicado: (2026)
Anytime Training with Schedule-Free Spectral Optimization
por: Apte, Anuj, et al.
Publicado: (2026)
por: Apte, Anuj, et al.
Publicado: (2026)
Constructing Industrial-Scale Optimization Modeling Benchmark
por: Li, Zhong, et al.
Publicado: (2026)
por: Li, Zhong, et al.
Publicado: (2026)
On Some Tunable Multi-fidelity Bayesian Optimization Frameworks
por: Manoj, Arjun, et al.
Publicado: (2025)
por: Manoj, Arjun, et al.
Publicado: (2025)
How Memory in Optimization Algorithms Implicitly Modifies the Loss
por: Cattaneo, Matias D., et al.
Publicado: (2025)
por: Cattaneo, Matias D., et al.
Publicado: (2025)
Ejemplares similares
-
Differentiable Distributionally Robust Optimization Layers
por: Ma, Xutao, et al.
Publicado: (2024) -
Preconditioning Benefits of Spectral Orthogonalization in Muon
por: Ma, Jianhao, et al.
Publicado: (2026) -
Uncovering Symmetry Transfer in Large Language Models via Layer-Peeled Optimization
por: Du, Zhehang, et al.
Publicado: (2026) -
PDE Control Gym: A Benchmark for Data-Driven Boundary Control of Partial Differential Equations
por: Bhan, Luke, et al.
Publicado: (2024) -
Differentiable Optimization for Deep Learning-Enhanced DC Approximation of AC Optimal Power Flow
por: Rosemberg, Andrew, et al.
Publicado: (2025)