Improving the Straight-Through Estimator with Zeroth-Order Information
Fuente:
arXiv
Salvato in:
| Autori principali: | Yang, Ningfeng, Aamodt, Tor M. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Boosting Entropy with Bell Box Quantization
di: Yang, Ningfeng, et al.
Pubblicazione: (2026)
di: Yang, Ningfeng, et al.
Pubblicazione: (2026)
ReFrame: Layer Caching for Accelerated Inference in Real-Time Rendering
di: Liu, Lufei, et al.
Pubblicazione: (2025)
di: Liu, Lufei, et al.
Pubblicazione: (2025)
Improving Discrete Optimisation Via Decoupled Straight-Through Estimator
di: Shah, Rushi, et al.
Pubblicazione: (2024)
di: Shah, Rushi, et al.
Pubblicazione: (2024)
Custom Gradient Estimators are Straight-Through Estimators in Disguise
di: Schoenbauer, Matt, et al.
Pubblicazione: (2024)
di: Schoenbauer, Matt, et al.
Pubblicazione: (2024)
Stochastic Dimension-Free Zeroth-Order Estimator for High-Dimensional and High-Order PINNs
di: Liang, Zhangyong, et al.
Pubblicazione: (2026)
di: Liang, Zhangyong, et al.
Pubblicazione: (2026)
PV-Tuning: Beyond Straight-Through Estimation for Extreme LLM Compression
di: Malinovskii, Vladimir, et al.
Pubblicazione: (2024)
di: Malinovskii, Vladimir, et al.
Pubblicazione: (2024)
On Adaptivity in Zeroth-Order Optimization
di: Dbouk, Hassan, et al.
Pubblicazione: (2026)
di: Dbouk, Hassan, et al.
Pubblicazione: (2026)
Data-Free Black-Box Federated Learning via Zeroth-Order Gradient Estimation
di: Ma, Xinge, et al.
Pubblicazione: (2025)
di: Ma, Xinge, et al.
Pubblicazione: (2025)
Learning Dynamics of Zeroth-Order Optimization: A Kernel Perspective
di: Li, Zhe, et al.
Pubblicazione: (2026)
di: Li, Zhe, et al.
Pubblicazione: (2026)
On the Optimal Construction of Unbiased Gradient Estimators for Zeroth-Order Optimization
di: Ma, Shaocong, et al.
Pubblicazione: (2025)
di: Ma, Shaocong, et al.
Pubblicazione: (2025)
Zeroth-Order Optimization is Secretly Single-Step Policy Optimization
di: Qiu, Junbin, et al.
Pubblicazione: (2025)
di: Qiu, Junbin, et al.
Pubblicazione: (2025)
Private Zeroth-Order Optimization with Public Data
di: Gong, Xuchen, et al.
Pubblicazione: (2025)
di: Gong, Xuchen, et al.
Pubblicazione: (2025)
The Multi-Query Paradox in Zeroth-Order Optimization
di: Lin, Wei, et al.
Pubblicazione: (2025)
di: Lin, Wei, et al.
Pubblicazione: (2025)
Fed-ZOE: Communication-Efficient Over-the-Air Federated Learning via Zeroth-Order Estimation
di: Jang, Jonggyu, et al.
Pubblicazione: (2024)
di: Jang, Jonggyu, et al.
Pubblicazione: (2024)
Elucidating Subspace Perturbation in Zeroth-Order Optimization: Theory and Practice at Scale
di: Park, Sihwan, et al.
Pubblicazione: (2025)
di: Park, Sihwan, et al.
Pubblicazione: (2025)
LLM Zeroth-Order Fine-Tuning is an Inference Workload
di: Li, Zelin, et al.
Pubblicazione: (2026)
di: Li, Zelin, et al.
Pubblicazione: (2026)
Alternating Diffusion for Proximal Sampling with Zeroth Order Queries
di: Takagi, Hirohane, et al.
Pubblicazione: (2026)
di: Takagi, Hirohane, et al.
Pubblicazione: (2026)
Riemannian Zeroth-Order Gradient Estimation with Structure-Preserving Metrics for Geodesically Incomplete Manifolds
di: Ma, Shaocong, et al.
Pubblicazione: (2026)
di: Ma, Shaocong, et al.
Pubblicazione: (2026)
Gradient Compressed Sensing: A Query-Efficient Gradient Estimator for High-Dimensional Zeroth-Order Optimization
di: Qiu, Ruizhong, et al.
Pubblicazione: (2024)
di: Qiu, Ruizhong, et al.
Pubblicazione: (2024)
Zeroth-Order Fine-Tuning of LLMs with Extreme Sparsity
di: Guo, Wentao, et al.
Pubblicazione: (2024)
di: Guo, Wentao, et al.
Pubblicazione: (2024)
Extending Straight-Through Estimation for Robust Neural Networks on Analog CIM Hardware
di: Feng, Yuannuo, et al.
Pubblicazione: (2025)
di: Feng, Yuannuo, et al.
Pubblicazione: (2025)
Refining Adaptive Zeroth-Order Optimization at Ease
di: Shu, Yao, et al.
Pubblicazione: (2025)
di: Shu, Yao, et al.
Pubblicazione: (2025)
Zeroth-Order Optimization at the Edge of Stability
di: Song, Minhak, et al.
Pubblicazione: (2026)
di: Song, Minhak, et al.
Pubblicazione: (2026)
Learning a Zeroth-Order Optimizer for Fine-Tuning LLMs
di: Zhang, Kairun, et al.
Pubblicazione: (2025)
di: Zhang, Kairun, et al.
Pubblicazione: (2025)
Z0-Inf: Zeroth Order Approximation for Data Influence
di: Kokhlikyan, Narine, et al.
Pubblicazione: (2025)
di: Kokhlikyan, Narine, et al.
Pubblicazione: (2025)
Position: Zeroth-Order Optimization in Deep Learning Is Underexplored, Not Underpowered
di: Liu, Sijia, et al.
Pubblicazione: (2026)
di: Liu, Sijia, et al.
Pubblicazione: (2026)
Efficient Federated RLHF via Zeroth-Order Policy Optimization
di: Wang, Deyi, et al.
Pubblicazione: (2026)
di: Wang, Deyi, et al.
Pubblicazione: (2026)
Compander-Aligned Query Geometry for Quantized Zeroth-Order Optimization
di: Shu, Yao, et al.
Pubblicazione: (2026)
di: Shu, Yao, et al.
Pubblicazione: (2026)
Zeroth-Order Sharpness-Aware Learning with Exponential Tilting
di: Gong, Xuchen, et al.
Pubblicazione: (2025)
di: Gong, Xuchen, et al.
Pubblicazione: (2025)
On the Inherent Privacy of Zeroth Order Projected Gradient Descent
di: Gupta, Devansh, et al.
Pubblicazione: (2025)
di: Gupta, Devansh, et al.
Pubblicazione: (2025)
Zeroth-Order Fine-Tuning of LLMs in Random Subspaces
di: Yu, Ziming, et al.
Pubblicazione: (2024)
di: Yu, Ziming, et al.
Pubblicazione: (2024)
Zeroth-Order Non-Log-Concave Sampling with Variance Reduction and Applications to Inverse Problems
di: Sahin, M. Berk, et al.
Pubblicazione: (2026)
di: Sahin, M. Berk, et al.
Pubblicazione: (2026)
Privacy Amplification in Differentially Private Zeroth-Order Optimization with Hidden States
di: Chien, Eli, et al.
Pubblicazione: (2025)
di: Chien, Eli, et al.
Pubblicazione: (2025)
Low-Rank Curvature for Zeroth-Order Optimization in LLM Fine-Tuning
di: Seung, Hyunseok, et al.
Pubblicazione: (2025)
di: Seung, Hyunseok, et al.
Pubblicazione: (2025)
Powering Up Zeroth-Order Training via Subspace Gradient Orthogonalization
di: Lang, Yicheng, et al.
Pubblicazione: (2026)
di: Lang, Yicheng, et al.
Pubblicazione: (2026)
High-Dimensional Learning Dynamics of Quantized Models with Straight-Through Estimator
di: Ichikawa, Yuma, et al.
Pubblicazione: (2025)
di: Ichikawa, Yuma, et al.
Pubblicazione: (2025)
Zeroth-Order Optimization Finds Flat Minima
di: Zhang, Liang, et al.
Pubblicazione: (2025)
di: Zhang, Liang, et al.
Pubblicazione: (2025)
Certified Multi-Fidelity Zeroth-Order Optimization
di: de Montbrun, Étienne, et al.
Pubblicazione: (2023)
di: de Montbrun, Étienne, et al.
Pubblicazione: (2023)
Private Zeroth-Order Nonsmooth Nonconvex Optimization
di: Zhang, Qinzi, et al.
Pubblicazione: (2024)
di: Zhang, Qinzi, et al.
Pubblicazione: (2024)
Second-Order Fine-Tuning without Pain for LLMs:A Hessian Informed Zeroth-Order Optimizer
di: Zhao, Yanjun, et al.
Pubblicazione: (2024)
di: Zhao, Yanjun, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Boosting Entropy with Bell Box Quantization
di: Yang, Ningfeng, et al.
Pubblicazione: (2026) -
ReFrame: Layer Caching for Accelerated Inference in Real-Time Rendering
di: Liu, Lufei, et al.
Pubblicazione: (2025) -
Improving Discrete Optimisation Via Decoupled Straight-Through Estimator
di: Shah, Rushi, et al.
Pubblicazione: (2024) -
Custom Gradient Estimators are Straight-Through Estimators in Disguise
di: Schoenbauer, Matt, et al.
Pubblicazione: (2024) -
Stochastic Dimension-Free Zeroth-Order Estimator for High-Dimensional and High-Order PINNs
di: Liang, Zhangyong, et al.
Pubblicazione: (2026)