Gespeichert in:
| Hauptverfasser: | Wang, Yue, Zhou, Yi, Zou, Shaofeng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2022
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2209.02555 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Non-Asymptotic Analysis for Single-Loop (Natural) Actor-Critic with Compatible Function Approximation
von: Wang, Yudan, et al.
Veröffentlicht: (2024)
von: Wang, Yudan, et al.
Veröffentlicht: (2024)
Achieving the Asymptotically Optimal Sample Complexity of Offline Reinforcement Learning: A DRO-Based Approach
von: Wang, Yue, et al.
Veröffentlicht: (2023)
von: Wang, Yue, et al.
Veröffentlicht: (2023)
Model-Free Robust Reinforcement Learning with Sample Complexity Analysis
von: Wang, Yudan, et al.
Veröffentlicht: (2024)
von: Wang, Yudan, et al.
Veröffentlicht: (2024)
Convergence Guarantees for RMSProp and Adam in Generalized-smooth Non-convex Optimization with Affine Noise Variance
von: Zhang, Qi, et al.
Veröffentlicht: (2024)
von: Zhang, Qi, et al.
Veröffentlicht: (2024)
GQ-VAE: A gated quantized VAE for learning variable length tokens
von: Datta, Theo, et al.
Veröffentlicht: (2025)
von: Datta, Theo, et al.
Veröffentlicht: (2025)
Large-Scale Non-convex Stochastic Constrained Distributionally Robust Optimization
von: Zhang, Qi, et al.
Veröffentlicht: (2024)
von: Zhang, Qi, et al.
Veröffentlicht: (2024)
Adaptive Gradient Normalization and Independent Sampling for (Stochastic) Generalized-Smooth Optimization
von: Yang, Yufeng, et al.
Veröffentlicht: (2024)
von: Yang, Yufeng, et al.
Veröffentlicht: (2024)
Lower Bounds for Greedy Teaching Set Constructions
von: Compton, Spencer, et al.
Veröffentlicht: (2025)
von: Compton, Spencer, et al.
Veröffentlicht: (2025)
Near-Optimal Sample Complexity for Iterated CVaR Reinforcement Learning with a Generative Model
von: Deng, Zilong, et al.
Veröffentlicht: (2025)
von: Deng, Zilong, et al.
Veröffentlicht: (2025)
Finite-Time Logarithmic Bayes Regret Upper Bounds
von: Atsidakou, Alexia, et al.
Veröffentlicht: (2023)
von: Atsidakou, Alexia, et al.
Veröffentlicht: (2023)
Why Adam Can Beat SGD: Second-Moment Normalization Yields Sharper Tails
von: Jin, Ruinan, et al.
Veröffentlicht: (2026)
von: Jin, Ruinan, et al.
Veröffentlicht: (2026)
Finite-Sample Wasserstein Error Bounds and Concentration Inequalities for Nonlinear Stochastic Approximation
von: Kong, Seo Taek, et al.
Veröffentlicht: (2026)
von: Kong, Seo Taek, et al.
Veröffentlicht: (2026)
Step-level Denoising-time Diffusion Alignment with Multiple Objectives
von: Zhang, Qi, et al.
Veröffentlicht: (2026)
von: Zhang, Qi, et al.
Veröffentlicht: (2026)
Finite Neural Networks as Mixtures of Gaussian Processes: From Provable Error Bounds to Prior Selection
von: Adams, Steven, et al.
Veröffentlicht: (2024)
von: Adams, Steven, et al.
Veröffentlicht: (2024)
Detector-Evasive LLM Paraphrasing via Constrained Policy Optimization
von: Wang, Mingyi, et al.
Veröffentlicht: (2026)
von: Wang, Mingyi, et al.
Veröffentlicht: (2026)
Finite-Time Bounds for Average-Reward Fitted Q-Iteration
von: Lee, Jongmin, et al.
Veröffentlicht: (2025)
von: Lee, Jongmin, et al.
Veröffentlicht: (2025)
LDC-MTL: Balancing Multi-Task Learning through Scalable Loss Discrepancy Control
von: Xiao, Peiyao, et al.
Veröffentlicht: (2025)
von: Xiao, Peiyao, et al.
Veröffentlicht: (2025)
Sample Complexity Characterization for Linear Contextual MDPs
von: Deng, Junze, et al.
Veröffentlicht: (2024)
von: Deng, Junze, et al.
Veröffentlicht: (2024)
Constrained Reinforcement Learning Under Model Mismatch
von: Sun, Zhongchang, et al.
Veröffentlicht: (2024)
von: Sun, Zhongchang, et al.
Veröffentlicht: (2024)
A Finite Sample Complexity Bound for Distributionally Robust Q-learning
von: Wang, Shengbo, et al.
Veröffentlicht: (2023)
von: Wang, Shengbo, et al.
Veröffentlicht: (2023)
Theoretical Study of Conflict-Avoidant Multi-Objective Reinforcement Learning
von: Wang, Yudan, et al.
Veröffentlicht: (2024)
von: Wang, Yudan, et al.
Veröffentlicht: (2024)
Lower Bound on the Greedy Approximation Ratio for Adaptive Submodular Cover
von: Harris, Blake, et al.
Veröffentlicht: (2024)
von: Harris, Blake, et al.
Veröffentlicht: (2024)
LINC: Decoupling Local Consequence Scoring from Hidden Matching in Constructive Neural Routing
von: Qin, Shaofeng, et al.
Veröffentlicht: (2026)
von: Qin, Shaofeng, et al.
Veröffentlicht: (2026)
Operator Learning for Schrödinger Equation: Unitarity, Error Bounds, and Time Generalization
von: Patel, Yash, et al.
Veröffentlicht: (2025)
von: Patel, Yash, et al.
Veröffentlicht: (2025)
Finite-Time Error Analysis of Soft Q-Learning: Switching System Approach
von: Jeong, Narim, et al.
Veröffentlicht: (2024)
von: Jeong, Narim, et al.
Veröffentlicht: (2024)
Error-quantified Conformal Inference for Time Series
von: Wu, Junxi, et al.
Veröffentlicht: (2025)
von: Wu, Junxi, et al.
Veröffentlicht: (2025)
The Exploration of Error Bounds in Classification with Noisy Labels
von: Liu, Haixia, et al.
Veröffentlicht: (2025)
von: Liu, Haixia, et al.
Veröffentlicht: (2025)
A Greedy Strategy for Graph Cut
von: Nie, Feiping, et al.
Veröffentlicht: (2024)
von: Nie, Feiping, et al.
Veröffentlicht: (2024)
MGDA Converges under Generalized Smoothness, Provably
von: Zhang, Qi, et al.
Veröffentlicht: (2024)
von: Zhang, Qi, et al.
Veröffentlicht: (2024)
Nonasymptotic CLT and Error Bounds for Two-Time-Scale Stochastic Approximation
von: Kong, Seo Taek, et al.
Veröffentlicht: (2025)
von: Kong, Seo Taek, et al.
Veröffentlicht: (2025)
Greedy Information Projection for LLM Data Selection
von: Dong, Victor Ye, et al.
Veröffentlicht: (2026)
von: Dong, Victor Ye, et al.
Veröffentlicht: (2026)
Relative Error Bound Analysis for Nuclear Norm Regularized Matrix Completion
von: Zhang, Lijun, et al.
Veröffentlicht: (2015)
von: Zhang, Lijun, et al.
Veröffentlicht: (2015)
Optimization of Epsilon-Greedy Exploration
von: Che, Ethan, et al.
Veröffentlicht: (2025)
von: Che, Ethan, et al.
Veröffentlicht: (2025)
Extremely Greedy Equivalence Search
von: Nazaret, Achille, et al.
Veröffentlicht: (2025)
von: Nazaret, Achille, et al.
Veröffentlicht: (2025)
Error Bounds for Flow Matching Methods
von: Benton, Joe, et al.
Veröffentlicht: (2023)
von: Benton, Joe, et al.
Veröffentlicht: (2023)
Classification Error Bound for Low Bayes Error Conditions in Machine Learning
von: Yang, Zijian, et al.
Veröffentlicht: (2025)
von: Yang, Zijian, et al.
Veröffentlicht: (2025)
Error Slice Discovery via Manifold Compactness
von: Yu, Han, et al.
Veröffentlicht: (2025)
von: Yu, Han, et al.
Veröffentlicht: (2025)
Greedy Alignment Principle for Optimizer Selection
von: Lee, Jaerin, et al.
Veröffentlicht: (2025)
von: Lee, Jaerin, et al.
Veröffentlicht: (2025)
QGFN: Controllable Greediness with Action Values
von: Lau, Elaine, et al.
Veröffentlicht: (2024)
von: Lau, Elaine, et al.
Veröffentlicht: (2024)
Revisiting Randomization in Greedy Model Search
von: Chen, Xin, et al.
Veröffentlicht: (2025)
von: Chen, Xin, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Non-Asymptotic Analysis for Single-Loop (Natural) Actor-Critic with Compatible Function Approximation
von: Wang, Yudan, et al.
Veröffentlicht: (2024) -
Achieving the Asymptotically Optimal Sample Complexity of Offline Reinforcement Learning: A DRO-Based Approach
von: Wang, Yue, et al.
Veröffentlicht: (2023) -
Model-Free Robust Reinforcement Learning with Sample Complexity Analysis
von: Wang, Yudan, et al.
Veröffentlicht: (2024) -
Convergence Guarantees for RMSProp and Adam in Generalized-smooth Non-convex Optimization with Affine Noise Variance
von: Zhang, Qi, et al.
Veröffentlicht: (2024) -
GQ-VAE: A gated quantized VAE for learning variable length tokens
von: Datta, Theo, et al.
Veröffentlicht: (2025)