Gespeichert in:
| Hauptverfasser: | Fang, Zijian, Liu, Zongkai, Yu, Chao, Hu, Chaohao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2501.00533 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
An Offline Adaptation Framework for Constrained Multi-Objective Reinforcement Learning
von: Lin, Qian, et al.
Veröffentlicht: (2024)
von: Lin, Qian, et al.
Veröffentlicht: (2024)
Policy-regularized Offline Multi-objective Reinforcement Learning
von: Lin, Qian, et al.
Veröffentlicht: (2024)
von: Lin, Qian, et al.
Veröffentlicht: (2024)
Iterative Minimax Games with Coupled Linear Constraints
von: Zhang, Huiling, et al.
Veröffentlicht: (2022)
von: Zhang, Huiling, et al.
Veröffentlicht: (2022)
Learning Word Embedding with Better Distance Weighting and Window Size Scheduling
von: Yang, Chaohao, et al.
Veröffentlicht: (2024)
von: Yang, Chaohao, et al.
Veröffentlicht: (2024)
Momentum Contrastive Learning with Enhanced Negative Sampling and Hard Negative Filtering
von: Hoang, Duy, et al.
Veröffentlicht: (2025)
von: Hoang, Duy, et al.
Veröffentlicht: (2025)
CPGD: Toward Stable Rule-based Reinforcement Learning for Language Models
von: Liu, Zongkai, et al.
Veröffentlicht: (2025)
von: Liu, Zongkai, et al.
Veröffentlicht: (2025)
GAGPO: Generalized Advantage Grouped Policy Optimization
von: Zhu, Siyuan, et al.
Veröffentlicht: (2026)
von: Zhu, Siyuan, et al.
Veröffentlicht: (2026)
Offline Multi-Agent Reinforcement Learning via In-Sample Sequential Policy Optimization
von: Liu, Zongkai, et al.
Veröffentlicht: (2024)
von: Liu, Zongkai, et al.
Veröffentlicht: (2024)
Fast Stochastic Policy Gradient: Negative Momentum for Reinforcement Learning
von: Zhang, Haobin, et al.
Veröffentlicht: (2024)
von: Zhang, Haobin, et al.
Veröffentlicht: (2024)
Minimax-Optimal Policy Regret in Partially Observable Markov Games
von: Arora, Raman
Veröffentlicht: (2026)
von: Arora, Raman
Veröffentlicht: (2026)
Guiding Diffusion Models with Reinforcement Learning for Stable Molecule Generation
von: Zhou, Zhijian, et al.
Veröffentlicht: (2025)
von: Zhou, Zhijian, et al.
Veröffentlicht: (2025)
Reinforcement Learning for Game-Theoretic Resource Allocation on Graphs
von: An, Zijian, et al.
Veröffentlicht: (2025)
von: An, Zijian, et al.
Veröffentlicht: (2025)
Negative Momentum for Convex-Concave Optimization
von: Shugart, Henry, et al.
Veröffentlicht: (2026)
von: Shugart, Henry, et al.
Veröffentlicht: (2026)
Minimax Statistical Estimation under Wasserstein Contamination
von: Chao, Patrick, et al.
Veröffentlicht: (2023)
von: Chao, Patrick, et al.
Veröffentlicht: (2023)
Language Agents with Reinforcement Learning for Strategic Play in the Werewolf Game
von: Xu, Zelai, et al.
Veröffentlicht: (2023)
von: Xu, Zelai, et al.
Veröffentlicht: (2023)
Deep SOR Minimax Q-learning for Two-player Zero-sum Game
von: Gautam, Saksham, et al.
Veröffentlicht: (2025)
von: Gautam, Saksham, et al.
Veröffentlicht: (2025)
Minimax Signal Detection in Sparse Additive Models
von: Kotekal, Subhodh, et al.
Veröffentlicht: (2023)
von: Kotekal, Subhodh, et al.
Veröffentlicht: (2023)
On Statistical Rates of Conditional Diffusion Transformers: Approximation, Estimation and Minimax Optimality
von: Hu, Jerry Yao-Chieh, et al.
Veröffentlicht: (2024)
von: Hu, Jerry Yao-Chieh, et al.
Veröffentlicht: (2024)
Minimax Rates for Hyperbolic Hierarchical Learning
von: Rawal, Divit, et al.
Veröffentlicht: (2026)
von: Rawal, Divit, et al.
Veröffentlicht: (2026)
Minimax Generalized Cross-Entropy
von: Bondugula, Kartheek, et al.
Veröffentlicht: (2026)
von: Bondugula, Kartheek, et al.
Veröffentlicht: (2026)
ParaFormer: A Generalized PageRank Graph Transformer for Graph Representation Learning
von: Yuan, Chaohao, et al.
Veröffentlicht: (2025)
von: Yuan, Chaohao, et al.
Veröffentlicht: (2025)
AdaPM: a Partial Momentum Algorithm for LLM Training
von: Zhang, Yimu, et al.
Veröffentlicht: (2025)
von: Zhang, Yimu, et al.
Veröffentlicht: (2025)
Dynamic Momentum Recalibration in Online Gradient Learning
von: Yao, Zhipeng, et al.
Veröffentlicht: (2026)
von: Yao, Zhipeng, et al.
Veröffentlicht: (2026)
Decouple before Integration: Test-time Synthesis of SFT and RLVR Task Vectors
von: Yuan, Chaohao, et al.
Veröffentlicht: (2026)
von: Yuan, Chaohao, et al.
Veröffentlicht: (2026)
ASD Classification on Dynamic Brain Connectome using Temporal Random Walk with Transformer-based Dynamic Network Embedding
von: Piriyasatit, Suchanuch, et al.
Veröffentlicht: (2025)
von: Piriyasatit, Suchanuch, et al.
Veröffentlicht: (2025)
Minimax Optimal Reinforcement Learning with Quasi-Optimism
von: Lee, Harin, et al.
Veröffentlicht: (2025)
von: Lee, Harin, et al.
Veröffentlicht: (2025)
Minimax Optimal Q Learning with Nearest Neighbors
von: Zhao, Puning, et al.
Veröffentlicht: (2023)
von: Zhao, Puning, et al.
Veröffentlicht: (2023)
Minimax Regret Learning for Data with Heterogeneous Subgroups
von: Mo, Weibin, et al.
Veröffentlicht: (2024)
von: Mo, Weibin, et al.
Veröffentlicht: (2024)
Automated discovery of symbolic laws governing skill acquisition from naturally occurring data
von: Liu, Sannyuya, et al.
Veröffentlicht: (2024)
von: Liu, Sannyuya, et al.
Veröffentlicht: (2024)
Momentum Approximation in Asynchronous Private Federated Learning
von: Yu, Tao, et al.
Veröffentlicht: (2024)
von: Yu, Tao, et al.
Veröffentlicht: (2024)
Learning What to Recommend: Minimax Optimal Simple Regret in Logistic Bandits
von: Liu, Shuai, et al.
Veröffentlicht: (2026)
von: Liu, Shuai, et al.
Veröffentlicht: (2026)
Corruption-Robust Linear Bandits: Minimax Optimality and Gap-Dependent Misspecification
von: Liu, Haolin, et al.
Veröffentlicht: (2024)
von: Liu, Haolin, et al.
Veröffentlicht: (2024)
A Multi-Step Minimax Q-learning Algorithm for Two-Player Zero-Sum Markov Games
von: R, Shreyas S, et al.
Veröffentlicht: (2024)
von: R, Shreyas S, et al.
Veröffentlicht: (2024)
Unifying Structural Proximity and Equivalence for Enhanced Dynamic Network Embedding
von: Piriyasatit, Suchanuch, et al.
Veröffentlicht: (2025)
von: Piriyasatit, Suchanuch, et al.
Veröffentlicht: (2025)
Minimax-Optimal Multi-Agent Robust Reinforcement Learning
von: Jiao, Yuchen, et al.
Veröffentlicht: (2024)
von: Jiao, Yuchen, et al.
Veröffentlicht: (2024)
Efficient Large-Scale Learning of Minimax Risk Classifiers
von: Bondugula, Kartheek, et al.
Veröffentlicht: (2025)
von: Bondugula, Kartheek, et al.
Veröffentlicht: (2025)
Momentum Further Constrains Sharpness at the Edge of Stochastic Stability
von: Andreyev, Arseniy, et al.
Veröffentlicht: (2026)
von: Andreyev, Arseniy, et al.
Veröffentlicht: (2026)
Towards Sharper Risk Bounds for Minimax Problems
von: Zhu, Bowei, et al.
Veröffentlicht: (2024)
von: Zhu, Bowei, et al.
Veröffentlicht: (2024)
Penalty-Based First-Order Methods for Bilevel Optimization with Minimax and Constrained Lower-Level Problems
von: Shen, Yiyang, et al.
Veröffentlicht: (2026)
von: Shen, Yiyang, et al.
Veröffentlicht: (2026)
Minimax Semiparametric Learning With Approximate Sparsity
von: Bradic, Jelena, et al.
Veröffentlicht: (2019)
von: Bradic, Jelena, et al.
Veröffentlicht: (2019)
Ähnliche Einträge
-
An Offline Adaptation Framework for Constrained Multi-Objective Reinforcement Learning
von: Lin, Qian, et al.
Veröffentlicht: (2024) -
Policy-regularized Offline Multi-objective Reinforcement Learning
von: Lin, Qian, et al.
Veröffentlicht: (2024) -
Iterative Minimax Games with Coupled Linear Constraints
von: Zhang, Huiling, et al.
Veröffentlicht: (2022) -
Learning Word Embedding with Better Distance Weighting and Window Size Scheduling
von: Yang, Chaohao, et al.
Veröffentlicht: (2024) -
Momentum Contrastive Learning with Enhanced Negative Sampling and Hard Negative Filtering
von: Hoang, Duy, et al.
Veröffentlicht: (2025)