DFWLayer: Differentiable Frank-Wolfe Optimization Layer
Fuente:
arXiv
Guardado en:
| Autores principales: | Liu, Zixuan, Liu, Liu, Wang, Xueqian, Zhao, Peilin |
|---|---|
| Formato: | Preprint |
| Publicado: |
2023
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
FW-Merging: Scaling Model Merging with Frank-Wolfe Optimization
por: Chen, Hao Mark, et al.
Publicado: (2025)
por: Chen, Hao Mark, et al.
Publicado: (2025)
Empirical evaluation of the Frank-Wolfe methods for constructing white-box adversarial attacks
por: Korotkova, Kristina, et al.
Publicado: (2025)
por: Korotkova, Kristina, et al.
Publicado: (2025)
Injecting Imbalance Sensitivity for Multi-Task Learning
por: Zhou, Zhipeng, et al.
Publicado: (2025)
por: Zhou, Zhipeng, et al.
Publicado: (2025)
HDT: Hierarchical Discrete Transformer for Multivariate Time Series Forecasting
por: Feng, Shibo, et al.
Publicado: (2025)
por: Feng, Shibo, et al.
Publicado: (2025)
Revisiting Plasticity in Visual Reinforcement Learning: Data, Modules and Training Stages
por: Ma, Guozheng, et al.
Publicado: (2023)
por: Ma, Guozheng, et al.
Publicado: (2023)
Robust Guided Diffusion for Offline Black-Box Optimization
por: Chen, Can Sam, et al.
Publicado: (2024)
por: Chen, Can Sam, et al.
Publicado: (2024)
DeepMF: Deep Motion Factorization for Closed-Loop Safety-Critical Driving Scenario Simulation
por: Li, Yizhe, et al.
Publicado: (2024)
por: Li, Yizhe, et al.
Publicado: (2024)
DP-FedAdamW: An Efficient Optimizer for Differentially Private Federated Large Models
por: Liu, Jin, et al.
Publicado: (2026)
por: Liu, Jin, et al.
Publicado: (2026)
Continual Optimization with Symmetry Teleportation for Multi-Task Learning
por: Zhou, Zhipeng, et al.
Publicado: (2025)
por: Zhou, Zhipeng, et al.
Publicado: (2025)
Multi-Task Vehicle Routing Solver via Mixture of Specialized Experts under State-Decomposable MDP
por: Pan, Yuxin, et al.
Publicado: (2025)
por: Pan, Yuxin, et al.
Publicado: (2025)
Probing Implicit Bias in Semi-gradient Q-learning: Visualizing the Effective Loss Landscapes via the Fokker--Planck Equation
por: Yin, Shuyu, et al.
Publicado: (2024)
por: Yin, Shuyu, et al.
Publicado: (2024)
DNAD: Differentiable Neural Architecture Distillation
por: Rao, Xuan, et al.
Publicado: (2025)
por: Rao, Xuan, et al.
Publicado: (2025)
Comparing Few to Rank Many: Active Human Preference Learning using Randomized Frank-Wolfe
por: Thekumparampil, Kiran Koshy, et al.
Publicado: (2024)
por: Thekumparampil, Kiran Koshy, et al.
Publicado: (2024)
Tree-Preconditioned Differentiable Optimization and Axioms as Layers
por: Liao, Yuexin
Publicado: (2025)
por: Liao, Yuexin
Publicado: (2025)
HELENE: Hessian Layer-wise Clipping and Gradient Annealing for Accelerating Fine-tuning LLM with Zeroth-order Optimization
por: Zhao, Huaqin, et al.
Publicado: (2024)
por: Zhao, Huaqin, et al.
Publicado: (2024)
Linear $Q$-Learning Does Not Diverge in $L^2$: Convergence Rates to a Bounded Set
por: Liu, Xinyu, et al.
Publicado: (2025)
por: Liu, Xinyu, et al.
Publicado: (2025)
What Makes Looped Transformers Perform Better Than Non-Recursive Ones
por: Gong, Zixuan, et al.
Publicado: (2025)
por: Gong, Zixuan, et al.
Publicado: (2025)
Gradient Flow Drifting: Generative Modeling via Wasserstein Gradient Flows of KDE-Approximated Divergences
por: Cao, Jiarui, et al.
Publicado: (2026)
por: Cao, Jiarui, et al.
Publicado: (2026)
Layer Specialization Underlying Compositional Reasoning in Transformers
por: Liu, Jing
Publicado: (2025)
por: Liu, Jing
Publicado: (2025)
Principled Data Selection for Alignment: The Hidden Risks of Difficult Examples
por: Gao, Chengqian, et al.
Publicado: (2025)
por: Gao, Chengqian, et al.
Publicado: (2025)
Reactivation: Empirical NTK Dynamics Under Task Shifts
por: Liu, Yuzhi, et al.
Publicado: (2025)
por: Liu, Yuzhi, et al.
Publicado: (2025)
It's All Connected: A Journey Through Test-Time Memorization, Attentional Bias, Retention, and Online Optimization
por: Behrouz, Ali, et al.
Publicado: (2025)
por: Behrouz, Ali, et al.
Publicado: (2025)
SEGNO: Generalizing Equivariant Graph Neural Networks with Physical Inductive Biases
por: Liu, Yang, et al.
Publicado: (2023)
por: Liu, Yang, et al.
Publicado: (2023)
TEON: Tensorized Orthonormalization Beyond Layer-Wise Muon for Large Language Model Pre-Training
por: Zhang, Ruijie, et al.
Publicado: (2026)
por: Zhang, Ruijie, et al.
Publicado: (2026)
Differentiable Distributionally Robust Optimization Layers
por: Ma, Xutao, et al.
Publicado: (2024)
por: Ma, Xutao, et al.
Publicado: (2024)
DAMBench: A Multi-Modal Benchmark for Deep Learning-based Atmospheric Data Assimilation
por: Wang, Hao, et al.
Publicado: (2025)
por: Wang, Hao, et al.
Publicado: (2025)
Finite Sample Analysis of Linear Temporal Difference Learning with Arbitrary Features
por: Xie, Zixuan, et al.
Publicado: (2025)
por: Xie, Zixuan, et al.
Publicado: (2025)
UACER: An Uncertainty-Adaptive Critic Ensemble Framework for Robust Adversarial Reinforcement Learning
por: Wu, Jiaxi, et al.
Publicado: (2025)
por: Wu, Jiaxi, et al.
Publicado: (2025)
Mastering Massive Multi-Task Reinforcement Learning via Mixture-of-Expert Decision Transformer
por: Kong, Yilun, et al.
Publicado: (2025)
por: Kong, Yilun, et al.
Publicado: (2025)
Golden Layers and Where to Find Them: Improved Knowledge Editing for Large Language Models Via Layer Gradient Analysis
por: Datta, Shrestha, et al.
Publicado: (2026)
por: Datta, Shrestha, et al.
Publicado: (2026)
FX-DARTS: Designing Topology-unconstrained Architectures with Differentiable Architecture Search and Entropy-based Super-network Shrinking
por: Rao, Xuan, et al.
Publicado: (2025)
por: Rao, Xuan, et al.
Publicado: (2025)
Wavelet Fourier Diffuser: Frequency-Aware Diffusion Model for Reinforcement Learning
por: Luo, Yifu, et al.
Publicado: (2025)
por: Luo, Yifu, et al.
Publicado: (2025)
FoMEMO: Towards Foundation Models for Expensive Multi-objective Optimization
por: Yao, Yiming, et al.
Publicado: (2025)
por: Yao, Yiming, et al.
Publicado: (2025)
Layer-Aware Influence for Online Data Valuation Estimation
por: Yang, Ziao, et al.
Publicado: (2025)
por: Yang, Ziao, et al.
Publicado: (2025)
An Improved Grey Wolf Optimization Algorithm for Heart Disease Prediction
por: Niu, Sihan, et al.
Publicado: (2024)
por: Niu, Sihan, et al.
Publicado: (2024)
Robust Policy Expansion for Offline-to-Online RL under Diverse Data Corruption
por: He, Longxiang, et al.
Publicado: (2025)
por: He, Longxiang, et al.
Publicado: (2025)
When Fewer Layers Break More Chains: Layer Pruning Harms Test-Time Scaling in LLMs
por: Wang, Keyu, et al.
Publicado: (2025)
por: Wang, Keyu, et al.
Publicado: (2025)
Rethinking KL Regularization in RLHF: From Value Estimation to Gradient Optimization
por: Liu, Kezhao, et al.
Publicado: (2025)
por: Liu, Kezhao, et al.
Publicado: (2025)
The Wolf Within: Covert Injection of Malice into MLLM Societies via an MLLM Operative
por: Tan, Zhen, et al.
Publicado: (2024)
por: Tan, Zhen, et al.
Publicado: (2024)
ECHO: Entropy-Confidence Hybrid Optimization for Test-Time Reinforcement Learning
por: Zhao, Chu, et al.
Publicado: (2026)
por: Zhao, Chu, et al.
Publicado: (2026)
Ejemplares similares
-
FW-Merging: Scaling Model Merging with Frank-Wolfe Optimization
por: Chen, Hao Mark, et al.
Publicado: (2025) -
Empirical evaluation of the Frank-Wolfe methods for constructing white-box adversarial attacks
por: Korotkova, Kristina, et al.
Publicado: (2025) -
Injecting Imbalance Sensitivity for Multi-Task Learning
por: Zhou, Zhipeng, et al.
Publicado: (2025) -
HDT: Hierarchical Discrete Transformer for Multivariate Time Series Forecasting
por: Feng, Shibo, et al.
Publicado: (2025) -
Revisiting Plasticity in Visual Reinforcement Learning: Data, Modules and Training Stages
por: Ma, Guozheng, et al.
Publicado: (2023)