Adaptive Robust Estimator for Multi-Agent Reinforcement Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Zhongyi, Tian, Wan, Chen, Jingyu, Huang, Kangyao, Zhang, Huiming, Yang, Hui, Ren, Tao, Jiang, Jinyang, Peng, Yijie, Ban, Yikun, Zhuang, Fuzhen |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Counterfactual Credit Policy Optimization for Multi-Agent Collaboration
by: Li, Zhongyi, et al.
Published: (2026)
by: Li, Zhongyi, et al.
Published: (2026)
Omni-Masked Gradient Descent: Memory-Efficient Optimization via Mask Traversal with Improved Convergence
by: Yang, Hui, et al.
Published: (2026)
by: Yang, Hui, et al.
Published: (2026)
Sharper Generalization Bounds for Transformer
by: Li, Yawen, et al.
Published: (2026)
by: Li, Yawen, et al.
Published: (2026)
Heterogeneous Agent Collaborative Reinforcement Learning
by: Zhang, Zhixia, et al.
Published: (2026)
by: Zhang, Zhixia, et al.
Published: (2026)
UniFAR: A Unified Facet-Aware Retrieval Framework for Scientific Documents
by: Dou, Zheng, et al.
Published: (2026)
by: Dou, Zheng, et al.
Published: (2026)
Closing the Loop: Coordinating Inventory and Recommendation via Deep Reinforcement Learning on Multiple Timescales
by: Jiang, Jinyang, et al.
Published: (2025)
by: Jiang, Jinyang, et al.
Published: (2025)
Weak-Driven Learning: How Weak Agents make Strong Agents Stronger
by: Chen, Zehao, et al.
Published: (2026)
by: Chen, Zehao, et al.
Published: (2026)
Contextual Rollout Bandits for Reinforcement Learning with Verifiable Rewards
by: Lu, Xiaodong, et al.
Published: (2026)
by: Lu, Xiaodong, et al.
Published: (2026)
Machine Learning-Assisted High-Dimensional Matrix Estimation
by: Tian, Wan, et al.
Published: (2026)
by: Tian, Wan, et al.
Published: (2026)
Deep Reinforcement Learning for Solving Management Problems: Towards A Large Management Mode
by: Jiang, Jinyang, et al.
Published: (2024)
by: Jiang, Jinyang, et al.
Published: (2024)
Dual-Agent Deep Reinforcement Learning for Dynamic Pricing and Replenishment
by: Zheng, Yi, et al.
Published: (2024)
by: Zheng, Yi, et al.
Published: (2024)
GCL-OT: Graph Contrastive Learning with Optimal Transport for Heterophilic Text-Attributed Graphs
by: Ren, Yating, et al.
Published: (2025)
by: Ren, Yating, et al.
Published: (2025)
FLeW: Facet-Level and Adaptive Weighted Representation Learning of Scientific Documents
by: Dou, Zheng, et al.
Published: (2025)
by: Dou, Zheng, et al.
Published: (2025)
FLOPS: Forward Learning with OPtimal Sampling
by: Ren, Tao, et al.
Published: (2024)
by: Ren, Tao, et al.
Published: (2024)
Adaptive Refinement Protocols for Distributed Distribution Estimation under $\ell^p$-Losses
by: Yuan, Deheng, et al.
Published: (2024)
by: Yuan, Deheng, et al.
Published: (2024)
Adaptive and Robust DBSCAN with Multi-agent Reinforcement Learning
by: Peng, Hao, et al.
Published: (2025)
by: Peng, Hao, et al.
Published: (2025)
RiskMiner: Discovering Formulaic Alphas via Risk Seeking Monte Carlo Tree Search
by: Ren, Tao, et al.
Published: (2024)
by: Ren, Tao, et al.
Published: (2024)
RiskPO: Risk-based Policy Optimization via Verifiable Reward for LLM Post-Training
by: Ren, Tao, et al.
Published: (2025)
by: Ren, Tao, et al.
Published: (2025)
Examining Agents' Bias Amplification versus Suppression in Multi-Agent Systems
by: Wu, Zejian Eric, et al.
Published: (2026)
by: Wu, Zejian Eric, et al.
Published: (2026)
Forward Learning with Differential Privacy
by: Feng, Mingqian, et al.
Published: (2025)
by: Feng, Mingqian, et al.
Published: (2025)
Learning a Distributed Hierarchical Locomotion Controller for Embodied Cooperation
by: Hong, Chuye, et al.
Published: (2024)
by: Hong, Chuye, et al.
Published: (2024)
Stochastic Approximation Methods for Distortion Risk Measure Optimization
by: Jiang, Jinyang, et al.
Published: (2025)
by: Jiang, Jinyang, et al.
Published: (2025)
CoNNect: Connectivity-Based Regularization for Structural Pruning
by: Franssen, Christian, et al.
Published: (2025)
by: Franssen, Christian, et al.
Published: (2025)
Policy Improvement Reinforcement Learning
by: Wang, Huaiyang, et al.
Published: (2026)
by: Wang, Huaiyang, et al.
Published: (2026)
Balancing the Reasoning Load: Difficulty-Differentiated Policy Optimization with Length Redistribution for Efficient and Robust Reinforcement Learning
by: Xia, Yinan, et al.
Published: (2026)
by: Xia, Yinan, et al.
Published: (2026)
Tacit Learning with Adaptive Information Selection for Cooperative Multi-Agent Reinforcement Learning
by: Liu, Lunjun, et al.
Published: (2024)
by: Liu, Lunjun, et al.
Published: (2024)
Federated Reasoning Distillation Framework with Model Learnability-Aware Data Allocation
by: Guo, Wei, et al.
Published: (2026)
by: Guo, Wei, et al.
Published: (2026)
LLM-ALSO: LLM-Driven Adaptive Learning-Signal Optimization for Multi-Agent Reinforcement Learning
by: Wu, Xiaoguang, et al.
Published: (2026)
by: Wu, Xiaoguang, et al.
Published: (2026)
Forward Learning for Gradient-based Black-box Saliency Map Generation
by: Zhang, Zeliang, et al.
Published: (2024)
by: Zhang, Zeliang, et al.
Published: (2024)
Distributed Nonparametric Estimation: from Sparse to Dense Samples per Terminal
by: Yuan, Deheng, et al.
Published: (2025)
by: Yuan, Deheng, et al.
Published: (2025)
Real-Time Aligned Reward Model beyond Semantics
by: Huang, Zixuan, et al.
Published: (2026)
by: Huang, Zixuan, et al.
Published: (2026)
Adaptive Multi-Agent Deep Reinforcement Learning for Timely Healthcare Interventions
by: Shaik, Thanveer, et al.
Published: (2023)
by: Shaik, Thanveer, et al.
Published: (2023)
When Self-Belief Misleads: Active Label Acquisition for Reinforcement Learning with Verifiable Rewards
by: Wang, Li, et al.
Published: (2026)
by: Wang, Li, et al.
Published: (2026)
Nonparametric Bayesian Optimization for General Rewards
by: Zhang, Zishi, et al.
Published: (2026)
by: Zhang, Zishi, et al.
Published: (2026)
One for Dozens: Adaptive REcommendation for All Domains with Counterfactual Augmentation
by: Luo, Huishi, et al.
Published: (2024)
by: Luo, Huishi, et al.
Published: (2024)
A Deep Reinforcement Learning Framework For Financial Portfolio Management
by: Li, Jinyang
Published: (2024)
by: Li, Jinyang
Published: (2024)
Hartle-Hawking state and its factorization in 3d gravity
by: Chua, Wan Zhen, et al.
Published: (2023)
by: Chua, Wan Zhen, et al.
Published: (2023)
Distributionally Robust Multi-Task Reinforcement Learning via Adaptive Task Sampling
by: Corrado, Nicholas E., et al.
Published: (2026)
by: Corrado, Nicholas E., et al.
Published: (2026)
Does Your Reasoning Model Implicitly Know When to Stop Thinking?
by: Huang, Zixuan, et al.
Published: (2026)
by: Huang, Zixuan, et al.
Published: (2026)
Causal Structure Representation Learning of Confounders in Latent Space for Recommendation
by: Xu, Hangtong, et al.
Published: (2023)
by: Xu, Hangtong, et al.
Published: (2023)
Similar Items
-
Counterfactual Credit Policy Optimization for Multi-Agent Collaboration
by: Li, Zhongyi, et al.
Published: (2026) -
Omni-Masked Gradient Descent: Memory-Efficient Optimization via Mask Traversal with Improved Convergence
by: Yang, Hui, et al.
Published: (2026) -
Sharper Generalization Bounds for Transformer
by: Li, Yawen, et al.
Published: (2026) -
Heterogeneous Agent Collaborative Reinforcement Learning
by: Zhang, Zhixia, et al.
Published: (2026) -
UniFAR: A Unified Facet-Aware Retrieval Framework for Scientific Documents
by: Dou, Zheng, et al.
Published: (2026)