Guardado en:
| Autores principales: | He, Yi, Zhou, Xingyu |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2605.07049 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
On the Sample Complexity of Differentially Private Policy Optimization
por: He, Yi, et al.
Publicado: (2025)
por: He, Yi, et al.
Publicado: (2025)
Improved Bounds for Private and Robust Alignment
por: Weng, Wenqian, et al.
Publicado: (2025)
por: Weng, Wenqian, et al.
Publicado: (2025)
Auditing Approximate Machine Unlearning for Differentially Private Models
por: Gu, Yuechun, et al.
Publicado: (2025)
por: Gu, Yuechun, et al.
Publicado: (2025)
On the Statistical Efficiency of Mean-Field Reinforcement Learning with General Function Approximation
por: Huang, Jiawei, et al.
Publicado: (2023)
por: Huang, Jiawei, et al.
Publicado: (2023)
Bellman Unbiasedness: Toward Provably Efficient Distributional Reinforcement Learning with General Value Function Approximation
por: Cho, Taehyun, et al.
Publicado: (2024)
por: Cho, Taehyun, et al.
Publicado: (2024)
Sample and Computationally Efficient Continuous-Time Reinforcement Learning with General Function Approximation
por: Zhao, Runze, et al.
Publicado: (2025)
por: Zhao, Runze, et al.
Publicado: (2025)
Differentially Private Deep Model-Based Reinforcement Learning
por: Rio, Alexandre, et al.
Publicado: (2024)
por: Rio, Alexandre, et al.
Publicado: (2024)
Towards User-level Private Reinforcement Learning with Human Feedback
por: Zhang, Jiaming, et al.
Publicado: (2025)
por: Zhang, Jiaming, et al.
Publicado: (2025)
Differentially Private Reinforcement Learning with Self-Play
por: Qiao, Dan, et al.
Publicado: (2024)
por: Qiao, Dan, et al.
Publicado: (2024)
A Unified Theoretical Analysis of Private and Robust Offline Alignment: from RLHF to DPO
por: Zhou, Xingyu, et al.
Publicado: (2025)
por: Zhou, Xingyu, et al.
Publicado: (2025)
Generalizing Differentially Private Decentralized Deep Learning with Multi-Agent Consensus
por: Bayrooti, Jasmine, et al.
Publicado: (2023)
por: Bayrooti, Jasmine, et al.
Publicado: (2023)
Efficient Differentially Private Fine-Tuning of LLMs via Reinforcement Learning
por: Khadangi, Afshin, et al.
Publicado: (2025)
por: Khadangi, Afshin, et al.
Publicado: (2025)
Tensor and Matrix Low-Rank Value-Function Approximation in Reinforcement Learning
por: Rozada, Sergio, et al.
Publicado: (2022)
por: Rozada, Sergio, et al.
Publicado: (2022)
The Role of Inherent Bellman Error in Offline Reinforcement Learning with Linear Function Approximation
por: Golowich, Noah, et al.
Publicado: (2024)
por: Golowich, Noah, et al.
Publicado: (2024)
Data-adaptive Differentially Private Prompt Synthesis for In-Context Learning
por: Gao, Fengyu, et al.
Publicado: (2024)
por: Gao, Fengyu, et al.
Publicado: (2024)
Differentially Private Worst-group Risk Minimization
por: Zhou, Xinyu, et al.
Publicado: (2024)
por: Zhou, Xinyu, et al.
Publicado: (2024)
Uncertainty-Aware Reward-Free Exploration with General Function Approximation
por: Zhang, Junkai, et al.
Publicado: (2024)
por: Zhang, Junkai, et al.
Publicado: (2024)
Kernel-Based Function Approximation for Average Reward Reinforcement Learning: An Optimist No-Regret Algorithm
por: Vakili, Sattar, et al.
Publicado: (2024)
por: Vakili, Sattar, et al.
Publicado: (2024)
Differentially Private Language Generation and Identification in the Limit
por: Mehrotra, Anay, et al.
Publicado: (2026)
por: Mehrotra, Anay, et al.
Publicado: (2026)
Learning with Differentially Private (Sliced) Wasserstein Gradients
por: Rodríguez-Vítores, David, et al.
Publicado: (2025)
por: Rodríguez-Vítores, David, et al.
Publicado: (2025)
Non-Asymptotic Analysis for Single-Loop (Natural) Actor-Critic with Compatible Function Approximation
por: Wang, Yudan, et al.
Publicado: (2024)
por: Wang, Yudan, et al.
Publicado: (2024)
Evaluating Differentially Private Generation of Domain-Specific Text
por: Sun, Yidan, et al.
Publicado: (2025)
por: Sun, Yidan, et al.
Publicado: (2025)
Function Approximation for Reinforcement Learning Controller for Energy from Spread Waves
por: Sarkar, Soumyendu, et al.
Publicado: (2024)
por: Sarkar, Soumyendu, et al.
Publicado: (2024)
DP-CSGP: Differentially Private Stochastic Gradient Push with Compressed Communication
por: Zhu, Zehan, et al.
Publicado: (2025)
por: Zhu, Zehan, et al.
Publicado: (2025)
Towards General-Purpose Model-Free Reinforcement Learning
por: Fujimoto, Scott, et al.
Publicado: (2025)
por: Fujimoto, Scott, et al.
Publicado: (2025)
Statistical Limits and Efficient Algorithms for Differentially Private Federated Learning
por: Auddy, Arnab, et al.
Publicado: (2026)
por: Auddy, Arnab, et al.
Publicado: (2026)
Neural Lyapunov Function Approximation with Self-Supervised Reinforcement Learning
por: McCutcheon, Luc, et al.
Publicado: (2025)
por: McCutcheon, Luc, et al.
Publicado: (2025)
DP-OPD: Differentially Private On-Policy Distillation for Language Models
por: Khadem, Fatemeh, et al.
Publicado: (2026)
por: Khadem, Fatemeh, et al.
Publicado: (2026)
Differentially Private Federated Clustering with Random Rebalancing
por: Yang, Xiyuan, et al.
Publicado: (2025)
por: Yang, Xiyuan, et al.
Publicado: (2025)
Multi-Objective Optimization for Privacy-Utility Balance in Differentially Private Federated Learning
por: Ranaweera, Kanishka, et al.
Publicado: (2025)
por: Ranaweera, Kanishka, et al.
Publicado: (2025)
Differentially Private In-Context Learning with Nearest Neighbor Search
por: Koskela, Antti, et al.
Publicado: (2025)
por: Koskela, Antti, et al.
Publicado: (2025)
RLCAD: Reinforcement Learning Training Gym for Revolution Involved CAD Command Sequence Generation
por: Yin, Xiaolong, et al.
Publicado: (2025)
por: Yin, Xiaolong, et al.
Publicado: (2025)
DP-FedPGN: Finding Global Flat Minima for Differentially Private Federated Learning via Penalizing Gradient Norm
por: Liu, Junkang, et al.
Publicado: (2025)
por: Liu, Junkang, et al.
Publicado: (2025)
Differentially Private Model Merging
por: Yin, Qichuan, et al.
Publicado: (2026)
por: Yin, Qichuan, et al.
Publicado: (2026)
Tackling Heavy-Tailed Rewards in Reinforcement Learning with Function Approximation: Minimax Optimal and Instance-Dependent Regret Bounds
por: Huang, Jiayi, et al.
Publicado: (2023)
por: Huang, Jiayi, et al.
Publicado: (2023)
Certification for Differentially Private Prediction in Gradient-Based Training
por: Wicker, Matthew, et al.
Publicado: (2024)
por: Wicker, Matthew, et al.
Publicado: (2024)
KL-regularization Itself is Differentially Private in Bandits and RLHF
por: Zhang, Yizhou, et al.
Publicado: (2025)
por: Zhang, Yizhou, et al.
Publicado: (2025)
Optimizing Automatic Differentiation with Deep Reinforcement Learning
por: Lohoff, Jamie, et al.
Publicado: (2024)
por: Lohoff, Jamie, et al.
Publicado: (2024)
A Differential Perspective on Distributional Reinforcement Learning
por: Rojas, Juan Sebastian, et al.
Publicado: (2025)
por: Rojas, Juan Sebastian, et al.
Publicado: (2025)
Overcoming Overfitting in Reinforcement Learning via Gaussian Process Diffusion Policy
por: Horprasert, Amornyos, et al.
Publicado: (2025)
por: Horprasert, Amornyos, et al.
Publicado: (2025)
Ejemplares similares
-
On the Sample Complexity of Differentially Private Policy Optimization
por: He, Yi, et al.
Publicado: (2025) -
Improved Bounds for Private and Robust Alignment
por: Weng, Wenqian, et al.
Publicado: (2025) -
Auditing Approximate Machine Unlearning for Differentially Private Models
por: Gu, Yuechun, et al.
Publicado: (2025) -
On the Statistical Efficiency of Mean-Field Reinforcement Learning with General Function Approximation
por: Huang, Jiawei, et al.
Publicado: (2023) -
Bellman Unbiasedness: Toward Provably Efficient Distributional Reinforcement Learning with General Value Function Approximation
por: Cho, Taehyun, et al.
Publicado: (2024)