Guardado en:
| Autores principales: | Hu, Tianmeng, Cui, Yongzheng, Luo, Biao, Li, Ke |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2602.16548 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Beyond Monotonicity: Revisiting Factorization Principles in Multi-Agent Q-Learning
por: Hu, Tianmeng, et al.
Publicado: (2025)
por: Hu, Tianmeng, et al.
Publicado: (2025)
PA2D-MORL: Pareto Ascent Directional Decomposition based Multi-Objective Reinforcement Learning
por: Hu, Tianmeng, et al.
Publicado: (2026)
por: Hu, Tianmeng, et al.
Publicado: (2026)
MO-MIX: Multi-Objective Multi-Agent Cooperative Decision-Making With Deep Reinforcement Learning
por: Hu, Tianmeng, et al.
Publicado: (2026)
por: Hu, Tianmeng, et al.
Publicado: (2026)
Uncertainty-Aware Multi-Objective Reinforcement Learning-Guided Diffusion Models for 3D De Novo Molecular Design
por: Chen, Lianghong, et al.
Publicado: (2025)
por: Chen, Lianghong, et al.
Publicado: (2025)
Inverse Design of Metamaterials with Manufacturing-Guiding Spectrum-to-Structure Conditional Diffusion Model
por: Li, Jiawen, et al.
Publicado: (2025)
por: Li, Jiawen, et al.
Publicado: (2025)
Environment Design for Inverse Reinforcement Learning
por: Buening, Thomas Kleine, et al.
Publicado: (2022)
por: Buening, Thomas Kleine, et al.
Publicado: (2022)
GenPO: Generative Diffusion Models Meet On-Policy Reinforcement Learning
por: Ding, Shutong, et al.
Publicado: (2025)
por: Ding, Shutong, et al.
Publicado: (2025)
Learning To Sample From Diffusion Models Via Inverse Reinforcement Learning
por: Bourdrez, Constant, et al.
Publicado: (2026)
por: Bourdrez, Constant, et al.
Publicado: (2026)
CausalGDP: Causality-Guided Diffusion Policies for Reinforcement Learning
por: Xiao, Xiaofeng, et al.
Publicado: (2026)
por: Xiao, Xiaofeng, et al.
Publicado: (2026)
Multi-Mode Process Control Using Multi-Task Inverse Reinforcement Learning
por: Lin, Runze, et al.
Publicado: (2025)
por: Lin, Runze, et al.
Publicado: (2025)
Distributional Reinforcement Learning with Diffusion Bridge Critics
por: Ding, Shutong, et al.
Publicado: (2026)
por: Ding, Shutong, et al.
Publicado: (2026)
Inverse Reinforcement Learning without Reinforcement Learning
por: Swamy, Gokul, et al.
Publicado: (2023)
por: Swamy, Gokul, et al.
Publicado: (2023)
Guided Diffusion for Fast Inverse Design of Density-based Mechanical Metamaterials
por: Yang, Yanyan, et al.
Publicado: (2024)
por: Yang, Yanyan, et al.
Publicado: (2024)
PAGAR: Taming Reward Misalignment in Inverse Reinforcement Learning-Based Imitation Learning with Protagonist Antagonist Guided Adversarial Reward
por: Zhou, Weichao, et al.
Publicado: (2023)
por: Zhou, Weichao, et al.
Publicado: (2023)
Distributional Inverse Reinforcement Learning
por: Wu, Feiyang, et al.
Publicado: (2025)
por: Wu, Feiyang, et al.
Publicado: (2025)
Prior-Guided Diffusion Planning for Offline Reinforcement Learning
por: Ki, Donghyeon, et al.
Publicado: (2025)
por: Ki, Donghyeon, et al.
Publicado: (2025)
Sample-Efficient Diffusion-based Reinforcement Learning with Critic Guidance
por: Ding, Shutong, et al.
Publicado: (2026)
por: Ding, Shutong, et al.
Publicado: (2026)
Inverse Design in Distributed Circuits Using Single-Step Reinforcement Learning
por: Li, Jiayu, et al.
Publicado: (2025)
por: Li, Jiayu, et al.
Publicado: (2025)
How Does the Lagrangian Guide Safe Reinforcement Learning through Diffusion Models?
por: Cheng, Xiaoyuan, et al.
Publicado: (2026)
por: Cheng, Xiaoyuan, et al.
Publicado: (2026)
Inverse Reinforcement Learning with Switching Rewards and History Dependency for Characterizing Animal Behaviors
por: Ke, Jingyang, et al.
Publicado: (2025)
por: Ke, Jingyang, et al.
Publicado: (2025)
Multi-Agent Reinforcement Learning for Inverse Design in Photonic Integrated Circuits
por: Mahlau, Yannik, et al.
Publicado: (2025)
por: Mahlau, Yannik, et al.
Publicado: (2025)
Safe Offline Reinforcement Learning with Feasibility-Guided Diffusion Model
por: Zheng, Yinan, et al.
Publicado: (2024)
por: Zheng, Yinan, et al.
Publicado: (2024)
Towards Generalized Inverse Reinforcement Learning
por: Dong, Chaosheng, et al.
Publicado: (2024)
por: Dong, Chaosheng, et al.
Publicado: (2024)
The Virtues of Pessimism in Inverse Reinforcement Learning
por: Wu, David, et al.
Publicado: (2024)
por: Wu, David, et al.
Publicado: (2024)
Energy-Guided Diffusion Sampling for Offline-to-Online Reinforcement Learning
por: Liu, Xu-Hui, et al.
Publicado: (2024)
por: Liu, Xu-Hui, et al.
Publicado: (2024)
Diffusion-based Reinforcement Learning via Q-weighted Variational Policy Optimization
por: Ding, Shutong, et al.
Publicado: (2024)
por: Ding, Shutong, et al.
Publicado: (2024)
Maximum Entropy Inverse Reinforcement Learning of Diffusion Models with Energy-Based Models
por: Yoon, Sangwoong, et al.
Publicado: (2024)
por: Yoon, Sangwoong, et al.
Publicado: (2024)
Hybrid Inverse Reinforcement Learning
por: Ren, Juntao, et al.
Publicado: (2024)
por: Ren, Juntao, et al.
Publicado: (2024)
A Bayesian Approach to Robust Inverse Reinforcement Learning
por: Wei, Ran, et al.
Publicado: (2023)
por: Wei, Ran, et al.
Publicado: (2023)
Advantage-Guided Diffusion for Model-Based Reinforcement Learning
por: Foffano, Daniele, et al.
Publicado: (2026)
por: Foffano, Daniele, et al.
Publicado: (2026)
Diffusion Guided Adversarial State Perturbations in Reinforcement Learning
por: Sun, Xiaolin, et al.
Publicado: (2025)
por: Sun, Xiaolin, et al.
Publicado: (2025)
RTLSeek: Boosting the LLM-Based RTL Generation with Multi-Stage Diversity-Oriented Reinforcement Learning
por: Zhang, Xinyu, et al.
Publicado: (2026)
por: Zhang, Xinyu, et al.
Publicado: (2026)
Scalable Multiagent Reinforcement Learning with Collective Influence Estimation
por: Luo, Zhenglong, et al.
Publicado: (2026)
por: Luo, Zhenglong, et al.
Publicado: (2026)
Kernel Density Bayesian Inverse Reinforcement Learning
por: Mandyam, Aishwarya, et al.
Publicado: (2023)
por: Mandyam, Aishwarya, et al.
Publicado: (2023)
Labeled TrustSet Guided: Batch Active Learning with Reinforcement Learning
por: Cui, Guofeng, et al.
Publicado: (2026)
por: Cui, Guofeng, et al.
Publicado: (2026)
Quantifying the Sensitivity of Inverse Reinforcement Learning to Misspecification
por: Skalse, Joar, et al.
Publicado: (2024)
por: Skalse, Joar, et al.
Publicado: (2024)
Inverse Reinforcement Learning with Multiple Planning Horizons
por: Yao, Jiayu, et al.
Publicado: (2024)
por: Yao, Jiayu, et al.
Publicado: (2024)
Accelerating Inverse Reinforcement Learning with Expert Bootstrapping
por: Wu, David, et al.
Publicado: (2024)
por: Wu, David, et al.
Publicado: (2024)
Walking the Values in Bayesian Inverse Reinforcement Learning
por: Bajgar, Ondrej, et al.
Publicado: (2024)
por: Bajgar, Ondrej, et al.
Publicado: (2024)
Confidence Aware Inverse Constrained Reinforcement Learning
por: Subramanian, Sriram Ganapathi, et al.
Publicado: (2024)
por: Subramanian, Sriram Ganapathi, et al.
Publicado: (2024)
Ejemplares similares
-
Beyond Monotonicity: Revisiting Factorization Principles in Multi-Agent Q-Learning
por: Hu, Tianmeng, et al.
Publicado: (2025) -
PA2D-MORL: Pareto Ascent Directional Decomposition based Multi-Objective Reinforcement Learning
por: Hu, Tianmeng, et al.
Publicado: (2026) -
MO-MIX: Multi-Objective Multi-Agent Cooperative Decision-Making With Deep Reinforcement Learning
por: Hu, Tianmeng, et al.
Publicado: (2026) -
Uncertainty-Aware Multi-Objective Reinforcement Learning-Guided Diffusion Models for 3D De Novo Molecular Design
por: Chen, Lianghong, et al.
Publicado: (2025) -
Inverse Design of Metamaterials with Manufacturing-Guiding Spectrum-to-Structure Conditional Diffusion Model
por: Li, Jiawen, et al.
Publicado: (2025)