Adversary-Robust Learning from Fully Asynchronous Directional Derivative Estimates
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Paul, Anik Kumar, Roy, Nibedita, Talagani, Nagesh, Ganesh, Swetha, Thoppe, Gugan, Reiffers-Masson, Alexandre |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Global Convergence Guarantees for Federated Policy Gradient Methods with Adversaries
von: Ganesh, Swetha, et al.
Veröffentlicht: (2024)
von: Ganesh, Swetha, et al.
Veröffentlicht: (2024)
Tight Convergence Rates for Online Distributed Linear Estimation with Adversarial Measurements
von: Roy, Nibedita, et al.
Veröffentlicht: (2026)
von: Roy, Nibedita, et al.
Veröffentlicht: (2026)
Does DQN Learn?
von: Gopalan, Aditya, et al.
Veröffentlicht: (2022)
von: Gopalan, Aditya, et al.
Veröffentlicht: (2022)
What Can Be Recovered Under Sparse Adversarial Corruption? Assumption-Free Theory for Linear Measurements
von: Halder, Vishal, et al.
Veröffentlicht: (2025)
von: Halder, Vishal, et al.
Veröffentlicht: (2025)
Online Learning of Weakly Coupled MDP Policies for Load Balancing and Auto Scaling
von: Eshwar, S. R., et al.
Veröffentlicht: (2024)
von: Eshwar, S. R., et al.
Veröffentlicht: (2024)
Reinforcement Learning with Quasi-Hyperbolic Discounting
von: Eshwar, S. R., et al.
Veröffentlicht: (2024)
von: Eshwar, S. R., et al.
Veröffentlicht: (2024)
Regularization for Adversarial Robust Learning
von: Wang, Jie, et al.
Veröffentlicht: (2024)
von: Wang, Jie, et al.
Veröffentlicht: (2024)
Stochastic Mirror Descent under Iterate-Dependent Markov Noise: Analysis in the Asymptotic and Finite Time Regimes
von: Paul, Anik Kumar, et al.
Veröffentlicht: (2026)
von: Paul, Anik Kumar, et al.
Veröffentlicht: (2026)
Zeroth-Order Non-smooth Non-convex Optimization via Gaussian Smoothing
von: Paul, Anik Kumar, et al.
Veröffentlicht: (2025)
von: Paul, Anik Kumar, et al.
Veröffentlicht: (2025)
Robust Analysis of Almost Sure Convergence of Zeroth-Order Mirror Descent Algorithm
von: Paul, Anik Kumar, et al.
Veröffentlicht: (2023)
von: Paul, Anik Kumar, et al.
Veröffentlicht: (2023)
Monotone and Conservative Policy Iteration Beyond the Tabular Case
von: Eshwar, S. R., et al.
Veröffentlicht: (2025)
von: Eshwar, S. R., et al.
Veröffentlicht: (2025)
Adversarially Robust Multitask Adaptive Control
von: Fallah, Kasra, et al.
Veröffentlicht: (2025)
von: Fallah, Kasra, et al.
Veröffentlicht: (2025)
Central Limit Theorems for Asynchronous Averaged Q-Learning
von: Liu, Xingtu
Veröffentlicht: (2025)
von: Liu, Xingtu
Veröffentlicht: (2025)
Fully Unconstrained Online Learning
von: Cutkosky, Ashok, et al.
Veröffentlicht: (2024)
von: Cutkosky, Ashok, et al.
Veröffentlicht: (2024)
Corruption-Tolerant Asynchronous Q-Learning with Near-Optimal Rates
von: Maity, Sreejeet, et al.
Veröffentlicht: (2025)
von: Maity, Sreejeet, et al.
Veröffentlicht: (2025)
Adversarially-Robust TD Learning with Markovian Data: Finite-Time Rates and Fundamental Limits
von: Maity, Sreejeet, et al.
Veröffentlicht: (2025)
von: Maity, Sreejeet, et al.
Veröffentlicht: (2025)
A Finite-Iteration Theory for Asynchronous Categorical Distributional Temporal-Difference Learning
von: Kaya, Ege C., et al.
Veröffentlicht: (2026)
von: Kaya, Ege C., et al.
Veröffentlicht: (2026)
Stochastic Recursive Inclusions under Biased Perturbations: An Input-to-State Stability Perspective
von: Paul, Anik Kumar, et al.
Veröffentlicht: (2026)
von: Paul, Anik Kumar, et al.
Veröffentlicht: (2026)
Asynchronous and Stochastic Distributed Resource Allocation
von: Li, Qiang, et al.
Veröffentlicht: (2025)
von: Li, Qiang, et al.
Veröffentlicht: (2025)
Self-Supervised Penalty-Based Learning for Robust Constrained Optimization
von: Benslimane, Wyame, et al.
Veröffentlicht: (2025)
von: Benslimane, Wyame, et al.
Veröffentlicht: (2025)
An Asynchronous Decentralised Optimisation Algorithm for Nonconvex Problems
von: Mafakheri, Behnam, et al.
Veröffentlicht: (2025)
von: Mafakheri, Behnam, et al.
Veröffentlicht: (2025)
Asynchronous Distributed Optimization with Delay-free Parameters
von: Wu, Xuyang, et al.
Veröffentlicht: (2023)
von: Wu, Xuyang, et al.
Veröffentlicht: (2023)
Distributionally and Adversarially Robust Logistic Regression via Intersecting Wasserstein Balls
von: Selvi, Aras, et al.
Veröffentlicht: (2024)
von: Selvi, Aras, et al.
Veröffentlicht: (2024)
Lower Bounds on Adversarial Robustness for Multiclass Classification with General Loss Functions
von: Trillos, Camilo Andrés García, et al.
Veröffentlicht: (2025)
von: Trillos, Camilo Andrés García, et al.
Veröffentlicht: (2025)
Adversarially and Distributionally Robust Virtual Energy Storage Systems via the Scenario Approach
von: Pantazis, Georgios, et al.
Veröffentlicht: (2025)
von: Pantazis, Georgios, et al.
Veröffentlicht: (2025)
Data-driven Reachable Set Estimation with Tunable Adversarial and Wasserstein Distributional Guarantees
von: Pantazis, Georgios, et al.
Veröffentlicht: (2026)
von: Pantazis, Georgios, et al.
Veröffentlicht: (2026)
Convergence Analysis of Stochastic Saddle Point Mirror Descent Algorithm -- A Projected Dynamical View Point
von: Paul, Anik Kumar, et al.
Veröffentlicht: (2024)
von: Paul, Anik Kumar, et al.
Veröffentlicht: (2024)
Almost Sure Convergence and Non-asymptotic Concentration Bounds for Stochastic Mirror Descent Algorithm
von: Paul, Anik Kumar, et al.
Veröffentlicht: (2024)
von: Paul, Anik Kumar, et al.
Veröffentlicht: (2024)
Blessings and Curses of Covariate Shifts: Adversarial Learning Dynamics, Directional Convergence, and Equilibria
von: Liang, Tengyuan
Veröffentlicht: (2022)
von: Liang, Tengyuan
Veröffentlicht: (2022)
Dual-Delayed Asynchronous SGD for Arbitrarily Heterogeneous Data
von: Wang, Xiaolu, et al.
Veröffentlicht: (2024)
von: Wang, Xiaolu, et al.
Veröffentlicht: (2024)
Encoder Decoder Generative Adversarial Network Model for Stock Market Prediction
von: Yadav, Bahadur, et al.
Veröffentlicht: (2025)
von: Yadav, Bahadur, et al.
Veröffentlicht: (2025)
Don't Explain Noise: Robust Counterfactuals for Randomized Ensembles
von: Forel, Alexandre, et al.
Veröffentlicht: (2022)
von: Forel, Alexandre, et al.
Veröffentlicht: (2022)
Robust Angular Synchronization via Directed Graph Neural Networks
von: He, Yixuan, et al.
Veröffentlicht: (2023)
von: He, Yixuan, et al.
Veröffentlicht: (2023)
Bayesian Ambiguity Contraction-based Adaptive Robust Markov Decision Processes for Adversarial Surveillance Missions
von: Choi, Jimin, et al.
Veröffentlicht: (2025)
von: Choi, Jimin, et al.
Veröffentlicht: (2025)
A Hessian-Free Actor-Critic Algorithm for Bi-Level Reinforcement Learning with Applications to LLM Fine-Tuning
von: Zeng, Sihan, et al.
Veröffentlicht: (2026)
von: Zeng, Sihan, et al.
Veröffentlicht: (2026)
Faster Stochastic Optimization with Arbitrary Delays via Asynchronous Mini-Batching
von: Attia, Amit, et al.
Veröffentlicht: (2024)
von: Attia, Amit, et al.
Veröffentlicht: (2024)
Optimal Asynchronous Stochastic Nonconvex Optimization under Heavy-Tailed Noise
von: Wu, Yidong, et al.
Veröffentlicht: (2026)
von: Wu, Yidong, et al.
Veröffentlicht: (2026)
Bilevel Models for Adversarial Learning and A Case Study
von: Zheng, Yutong, et al.
Veröffentlicht: (2025)
von: Zheng, Yutong, et al.
Veröffentlicht: (2025)
Early Directional Convergence in Deep Homogeneous Neural Networks for Small Initializations
von: Kumar, Akshay, et al.
Veröffentlicht: (2024)
von: Kumar, Akshay, et al.
Veröffentlicht: (2024)
Revisiting Multi-Agent Asynchronous Online Optimization with Delays: the Strongly Convex Case
von: Bao, Lingchan, et al.
Veröffentlicht: (2025)
von: Bao, Lingchan, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Global Convergence Guarantees for Federated Policy Gradient Methods with Adversaries
von: Ganesh, Swetha, et al.
Veröffentlicht: (2024) -
Tight Convergence Rates for Online Distributed Linear Estimation with Adversarial Measurements
von: Roy, Nibedita, et al.
Veröffentlicht: (2026) -
Does DQN Learn?
von: Gopalan, Aditya, et al.
Veröffentlicht: (2022) -
What Can Be Recovered Under Sparse Adversarial Corruption? Assumption-Free Theory for Linear Measurements
von: Halder, Vishal, et al.
Veröffentlicht: (2025) -
Online Learning of Weakly Coupled MDP Policies for Load Balancing and Auto Scaling
von: Eshwar, S. R., et al.
Veröffentlicht: (2024)