Saved in:
| Main Authors: | Yang, Yan, Gao, Bin, Yuan, Ya-xiang |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2501.13830 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Bilevel reinforcement learning via the development of hyper-gradient without lower-level convexity
by: Yang, Yan, et al.
Published: (2024)
by: Yang, Yan, et al.
Published: (2024)
LancBiO: dynamic Lanczos-aided bilevel optimization via Krylov subspace
by: Yang, Yan, et al.
Published: (2024)
by: Yang, Yan, et al.
Published: (2024)
Variational analysis of determinantal varieties
by: Yang, Yan, et al.
Published: (2025)
by: Yang, Yan, et al.
Published: (2025)
Efficient and provably convergent end-to-end training of deep neural networks with linear constraints
by: Yang, Zonglin, et al.
Published: (2026)
by: Yang, Zonglin, et al.
Published: (2026)
Stabilizing reinforcement learning control: A modular framework for optimizing over all stable behavior
by: Lawrence, Nathan P., et al.
Published: (2023)
by: Lawrence, Nathan P., et al.
Published: (2023)
Linear convergence of forward-backward accelerated algorithms without knowledge of the modulus of strong convexity
by: Li, Bowen, et al.
Published: (2023)
by: Li, Bowen, et al.
Published: (2023)
Weighted Low-rank Approximation via Stochastic Gradient Descent on Manifolds
by: Xu, Conglong, et al.
Published: (2025)
by: Xu, Conglong, et al.
Published: (2025)
Pinet: Optimizing hard-constrained neural networks with orthogonal projection layers
by: Grontas, Panagiotis D., et al.
Published: (2025)
by: Grontas, Panagiotis D., et al.
Published: (2025)
Sinkhorn doubly stochastic attention rank decay analysis
by: Lapenna, Michela, et al.
Published: (2026)
by: Lapenna, Michela, et al.
Published: (2026)
Optimization over the intersection of manifolds
by: Yang, Yan, et al.
Published: (2026)
by: Yang, Yan, et al.
Published: (2026)
Tutorial on amortized optimization
by: Amos, Brandon
Published: (2022)
by: Amos, Brandon
Published: (2022)
Towards Faster Decentralized Stochastic Optimization with Communication Compression
by: Islamov, Rustem, et al.
Published: (2024)
by: Islamov, Rustem, et al.
Published: (2024)
A second-order-like optimizer with adaptive gradient scaling for deep learning
by: Bolte, Jérôme, et al.
Published: (2024)
by: Bolte, Jérôme, et al.
Published: (2024)
Fuzzy hyperparameters update in a second order optimization
by: Bensadok, Abdelaziz, et al.
Published: (2024)
by: Bensadok, Abdelaziz, et al.
Published: (2024)
Frugality in second-order optimization: floating-point approximations for Newton's method
by: Carrino, Giuseppe, et al.
Published: (2025)
by: Carrino, Giuseppe, et al.
Published: (2025)
Linear attention is (maybe) all you need (to understand transformer optimization)
by: Ahn, Kwangjun, et al.
Published: (2023)
by: Ahn, Kwangjun, et al.
Published: (2023)
Desingularization of bounded-rank tensor sets
by: Gao, Bin, et al.
Published: (2024)
by: Gao, Bin, et al.
Published: (2024)
Accelerating RLHF Training with Reward Variance Increase
by: Yang, Zonglin, et al.
Published: (2025)
by: Yang, Zonglin, et al.
Published: (2025)
Deep learning enhanced mixed integer optimization: Learning to reduce model dimensionality
by: Triantafyllou, Niki, et al.
Published: (2024)
by: Triantafyllou, Niki, et al.
Published: (2024)
A second-order method landing on the Stiefel manifold via Newton$\unicode{x2013}$Schulz iteration
by: Xiong, Xinhui, et al.
Published: (2026)
by: Xiong, Xinhui, et al.
Published: (2026)
Low-rank optimization on Tucker tensor varieties
by: Gao, Bin, et al.
Published: (2023)
by: Gao, Bin, et al.
Published: (2023)
Double Momentum Method for Lower-Level Constrained Bilevel Optimization
by: Shi, Wanli, et al.
Published: (2024)
by: Shi, Wanli, et al.
Published: (2024)
High-dimensional mixed-categorical Gaussian processes with application to multidisciplinary design optimization for a green aircraft
by: Saves, Paul, et al.
Published: (2023)
by: Saves, Paul, et al.
Published: (2023)
A Theoretical Framework for Auxiliary-Loss-Free Load Balancing of Sparse Mixture-of-Experts in Large-Scale AI Models
by: Han, X. Y., et al.
Published: (2025)
by: Han, X. Y., et al.
Published: (2025)
Bridging Control with Neural Network Verifier alpha-beta-CROWN: A Tutorial
by: Li, Haoyu, et al.
Published: (2026)
by: Li, Haoyu, et al.
Published: (2026)
Benchmarking PtO and PnO Methods in the Predictive Combinatorial Optimization Regime
by: Geng, Haoyu, et al.
Published: (2023)
by: Geng, Haoyu, et al.
Published: (2023)
Scaling physics-informed hard constraints with mixture-of-experts
by: Chalapathi, Nithin, et al.
Published: (2024)
by: Chalapathi, Nithin, et al.
Published: (2024)
One-Layer Transformer Provably Learns One-Nearest Neighbor In Context
by: Li, Zihao, et al.
Published: (2024)
by: Li, Zihao, et al.
Published: (2024)
Multi-Year Maintenance Planning for Large-Scale Infrastructure Systems: A Novel Network Deep Q-Learning Approach
by: Fard, Amir, et al.
Published: (2025)
by: Fard, Amir, et al.
Published: (2025)
GANQ: GPU-Adaptive Non-Uniform Quantization for Large Language Models
by: Zhao, Pengxiang, et al.
Published: (2025)
by: Zhao, Pengxiang, et al.
Published: (2025)
SPAP: Structured Pruning via Alternating Optimization and Penalty Methods
by: Hu, Hanyu, et al.
Published: (2025)
by: Hu, Hanyu, et al.
Published: (2025)
A Deep Q-Network Based on Radial Basis Functions for Multi-Echelon Inventory Management
by: Cheng, Liqiang, et al.
Published: (2024)
by: Cheng, Liqiang, et al.
Published: (2024)
Capabilities of Large Language Models in Control Engineering: A Benchmark Study on GPT-4, Claude 3 Opus, and Gemini 1.0 Ultra
by: Kevian, Darioush, et al.
Published: (2024)
by: Kevian, Darioush, et al.
Published: (2024)
Deep Reinforcement Learning for Traveling Purchaser Problems
by: Yuan, Haofeng, et al.
Published: (2024)
by: Yuan, Haofeng, et al.
Published: (2024)
On the Convergence of (Stochastic) Gradient Descent for Kolmogorov--Arnold Networks
by: Gao, Yihang, et al.
Published: (2024)
by: Gao, Yihang, et al.
Published: (2024)
Reward-Directed Score-Based Diffusion Models via q-Learning
by: Gao, Xuefeng, et al.
Published: (2024)
by: Gao, Xuefeng, et al.
Published: (2024)
Hierarchical Deep Reinforcement Learning Framework for Multi-Year Asset Management Under Budget Constraints
by: Fard, Amir, et al.
Published: (2025)
by: Fard, Amir, et al.
Published: (2025)
Employing Deep Neural Operators for PDE control by decoupling training and optimization
by: Lundqvist, Oliver G. S., et al.
Published: (2025)
by: Lundqvist, Oliver G. S., et al.
Published: (2025)
Double-Bounded Optimal Transport for Advanced Clustering and Classification
by: Shi, Liangliang, et al.
Published: (2024)
by: Shi, Liangliang, et al.
Published: (2024)
Neural Solver Selection for Combinatorial Optimization
by: Gao, Chengrui, et al.
Published: (2024)
by: Gao, Chengrui, et al.
Published: (2024)
Similar Items
-
Bilevel reinforcement learning via the development of hyper-gradient without lower-level convexity
by: Yang, Yan, et al.
Published: (2024) -
LancBiO: dynamic Lanczos-aided bilevel optimization via Krylov subspace
by: Yang, Yan, et al.
Published: (2024) -
Variational analysis of determinantal varieties
by: Yang, Yan, et al.
Published: (2025) -
Efficient and provably convergent end-to-end training of deep neural networks with linear constraints
by: Yang, Zonglin, et al.
Published: (2026) -
Stabilizing reinforcement learning control: A modular framework for optimizing over all stable behavior
by: Lawrence, Nathan P., et al.
Published: (2023)