Guardado en:
| Autores principales: | Tang, Zhiwei, Rybin, Dmitry, Chang, Tsung-Hui |
|---|---|
| Formato: | Preprint |
| Publicado: |
2023
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2303.03751 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
FedLion: Faster Adaptive Federated Optimization with Fewer Communication
por: Tang, Zhiwei, et al.
Publicado: (2024)
por: Tang, Zhiwei, et al.
Publicado: (2024)
Provable Multi-Party Reinforcement Learning with Diverse Human Feedback
por: Zhong, Huiying, et al.
Publicado: (2024)
por: Zhong, Huiying, et al.
Publicado: (2024)
Inference-Time Alignment of Diffusion Models with Direct Noise Optimization
por: Tang, Zhiwei, et al.
Publicado: (2024)
por: Tang, Zhiwei, et al.
Publicado: (2024)
Minimisation of Quasar-Convex Functions Using Random Zeroth-Order Oracles
por: Farzin, Amir Ali, et al.
Publicado: (2025)
por: Farzin, Amir Ali, et al.
Publicado: (2025)
Refining Adaptive Zeroth-Order Optimization at Ease
por: Shu, Yao, et al.
Publicado: (2025)
por: Shu, Yao, et al.
Publicado: (2025)
A Historical Trajectory Assisted Optimization Method for Zeroth-Order Federated Learning
por: Wu, Chenlin, et al.
Publicado: (2024)
por: Wu, Chenlin, et al.
Publicado: (2024)
When Foresight Pruning Meets Zeroth-Order Optimization: Efficient Federated Learning for Low-Memory Devices
por: Zhang, Pengyu, et al.
Publicado: (2024)
por: Zhang, Pengyu, et al.
Publicado: (2024)
Zeroth-Order Optimization Finds Flat Minima
por: Zhang, Liang, et al.
Publicado: (2025)
por: Zhang, Liang, et al.
Publicado: (2025)
Subspace-based Approximate Hessian Method for Zeroth-Order Optimization
por: Kim, Dongyoon, et al.
Publicado: (2025)
por: Kim, Dongyoon, et al.
Publicado: (2025)
CurvZO: Adaptive Curvature-Guided Sparse Zeroth-Order Optimization for Efficient LLM Fine-Tuning
por: Wang, Shuo, et al.
Publicado: (2026)
por: Wang, Shuo, et al.
Publicado: (2026)
Robust and Efficient Zeroth-Order LLM Fine-Tuning via Adaptive Bayesian Subspace Optimizer
por: Feng, Jian, et al.
Publicado: (2026)
por: Feng, Jian, et al.
Publicado: (2026)
AR1-ZO: Topology-Aware Rank-1 Zeroth-Order Queries for High-Rank LoRA Fine-Tuning
por: Chen, Ziye, et al.
Publicado: (2026)
por: Chen, Ziye, et al.
Publicado: (2026)
On the Optimal Construction of Unbiased Gradient Estimators for Zeroth-Order Optimization
por: Ma, Shaocong, et al.
Publicado: (2025)
por: Ma, Shaocong, et al.
Publicado: (2025)
Scalable Valuation of Human Feedback through Provably Robust Model Alignment
por: Fujisawa, Masahiro, et al.
Publicado: (2025)
por: Fujisawa, Masahiro, et al.
Publicado: (2025)
Zeroth-Order Fine-Tuning of LLMs in Random Subspaces
por: Yu, Ziming, et al.
Publicado: (2024)
por: Yu, Ziming, et al.
Publicado: (2024)
Zeroth-Order Fine-Tuning of LLMs with Extreme Sparsity
por: Guo, Wentao, et al.
Publicado: (2024)
por: Guo, Wentao, et al.
Publicado: (2024)
Towards Off-Policy Reinforcement Learning for Ranking Policies with Human Feedback
por: Xiao, Teng, et al.
Publicado: (2024)
por: Xiao, Teng, et al.
Publicado: (2024)
OAT-Rephrase: Optimization-Aware Training Data Rephrasing for Zeroth-Order LLM Fine-Tuning
por: Long, Jikai, et al.
Publicado: (2025)
por: Long, Jikai, et al.
Publicado: (2025)
Implicit Hypergraph Neural Networks: A Stable Framework for Higher-Order Relational Learning with Provable Guarantees
por: Li, Xiaoyu, et al.
Publicado: (2025)
por: Li, Xiaoyu, et al.
Publicado: (2025)
Provable Interactive Learning with Hindsight Instruction Feedback
por: Misra, Dipendra, et al.
Publicado: (2024)
por: Misra, Dipendra, et al.
Publicado: (2024)
MaZO: Masked Zeroth-Order Optimization for Multi-Task Fine-Tuning of Large Language Models
por: Zhang, Zhen, et al.
Publicado: (2025)
por: Zhang, Zhen, et al.
Publicado: (2025)
FZOO: Fast Zeroth-Order Optimizer for Fine-Tuning Large Language Models towards Adam-Scale Speed
por: Dang, Sizhe, et al.
Publicado: (2025)
por: Dang, Sizhe, et al.
Publicado: (2025)
AdaMeZO: Adam-style Zeroth-Order Optimizer for LLM Fine-tuning Without Maintaining the Moments
por: Cai, Zhijie, et al.
Publicado: (2026)
por: Cai, Zhijie, et al.
Publicado: (2026)
Perturbation-efficient Zeroth-order Optimization for Hardware-friendly On-device Training
por: Tan, Qitao, et al.
Publicado: (2025)
por: Tan, Qitao, et al.
Publicado: (2025)
Prompt Optimization with Human Feedback
por: Lin, Xiaoqiang, et al.
Publicado: (2024)
por: Lin, Xiaoqiang, et al.
Publicado: (2024)
Communication-Efficient and Differentially Private Vertical Federated Learning with Zeroth-Order Optimization
por: Zhang, Jianing, et al.
Publicado: (2025)
por: Zhang, Jianing, et al.
Publicado: (2025)
Turning Stale Gradients into Stable Gradients: Coherent Coordinate Descent with Implicit Landscape Smoothing for Lightweight Zeroth-Order Optimization
por: Liang, Chen, et al.
Publicado: (2026)
por: Liang, Chen, et al.
Publicado: (2026)
Warming Up for Zeroth-Order Federated Pre-Training with Low Resource Clients
por: Legate, Gwen, et al.
Publicado: (2025)
por: Legate, Gwen, et al.
Publicado: (2025)
QuZO: Quantized Zeroth-Order Fine-Tuning for Large Language Models
por: Zhou, Jiajun, et al.
Publicado: (2025)
por: Zhou, Jiajun, et al.
Publicado: (2025)
Revisiting Zeroth-Order Optimization: Minimum-Variance Two-Point Estimators and Directionally Aligned Perturbations
por: Ma, Shaocong, et al.
Publicado: (2025)
por: Ma, Shaocong, et al.
Publicado: (2025)
Provably Learning from Modern Language Models via Low Logit Rank
por: Golowich, Noah, et al.
Publicado: (2025)
por: Golowich, Noah, et al.
Publicado: (2025)
Efficient Knowledge Graph Unlearning with Zeroth-order Information
por: Xiao, Yang, et al.
Publicado: (2025)
por: Xiao, Yang, et al.
Publicado: (2025)
Asynchronous Distributed Reinforcement Learning for LQR Control via Zeroth-Order Block Coordinate Descent
por: Jing, Gangshan, et al.
Publicado: (2021)
por: Jing, Gangshan, et al.
Publicado: (2021)
Adaptive Kernel Design for Bayesian Optimization Is a Piece of CAKE with LLMs
por: Suwandi, Richard Cornelius, et al.
Publicado: (2025)
por: Suwandi, Richard Cornelius, et al.
Publicado: (2025)
Enhancing Safety in Reinforcement Learning with Human Feedback via Rectified Policy Optimization
por: Peng, Xiyue, et al.
Publicado: (2024)
por: Peng, Xiyue, et al.
Publicado: (2024)
Mitigating Non-IID Drift in Zeroth-Order Federated LLM Fine-Tuning with Transferable Sparsity
por: Ran, Yide, et al.
Publicado: (2025)
por: Ran, Yide, et al.
Publicado: (2025)
Zeroth-Order Actor-Critic: An Evolutionary Framework for Sequential Decision Problems
por: Lei, Yuheng, et al.
Publicado: (2022)
por: Lei, Yuheng, et al.
Publicado: (2022)
Provable Privacy Advantages of Decentralized Federated Learning via Distributed Optimization
por: Yu, Wenrui, et al.
Publicado: (2024)
por: Yu, Wenrui, et al.
Publicado: (2024)
Provably Efficient Reinforcement Learning for Adversarial Restless Multi-Armed Bandits with Unknown Transitions and Bandit Feedback
por: Xiong, Guojun, et al.
Publicado: (2024)
por: Xiong, Guojun, et al.
Publicado: (2024)
Zeroth-Order Adaptive Neuron Alignment Based Pruning without Re-Training
por: Cunegatti, Elia, et al.
Publicado: (2024)
por: Cunegatti, Elia, et al.
Publicado: (2024)
Ejemplares similares
-
FedLion: Faster Adaptive Federated Optimization with Fewer Communication
por: Tang, Zhiwei, et al.
Publicado: (2024) -
Provable Multi-Party Reinforcement Learning with Diverse Human Feedback
por: Zhong, Huiying, et al.
Publicado: (2024) -
Inference-Time Alignment of Diffusion Models with Direct Noise Optimization
por: Tang, Zhiwei, et al.
Publicado: (2024) -
Minimisation of Quasar-Convex Functions Using Random Zeroth-Order Oracles
por: Farzin, Amir Ali, et al.
Publicado: (2025) -
Refining Adaptive Zeroth-Order Optimization at Ease
por: Shu, Yao, et al.
Publicado: (2025)