Imitating from auxiliary imperfect demonstrations via Adversarial Density Weighted Regression
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Ziqi, Zhuang, Zifeng, Xu, Jingzehua, Yang, Yiyuan, Huang, Yubo, Wang, Donglin, Zhang, Shuai |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Context-Former: Stitching via Latent Conditioned Sequence Modeling
by: Zhang, Ziqi, et al.
Published: (2024)
by: Zhang, Ziqi, et al.
Published: (2024)
A dynamical clipping approach with task feedback for Proximal Policy Optimization
by: Zhang, Ziqi, et al.
Published: (2023)
by: Zhang, Ziqi, et al.
Published: (2023)
Improving Offline-to-Online Reinforcement Learning with Q Conditioned State Entropy Exploration
by: Zhang, Ziqi, et al.
Published: (2023)
by: Zhang, Ziqi, et al.
Published: (2023)
Language Models as Efficient Reward Function Searchers for Custom-Environment Multi-Objective Reinforcement
by: Xie, Guanwen, et al.
Published: (2024)
by: Xie, Guanwen, et al.
Published: (2024)
Self-Training with Dynamic Weighting for Robust Gradual Domain Adaptation
by: Wang, Zixi, et al.
Published: (2025)
by: Wang, Zixi, et al.
Published: (2025)
TDMPBC: Self-Imitative Reinforcement Learning for Humanoid Robot Control
by: Zhuang, Zifeng, et al.
Published: (2025)
by: Zhuang, Zifeng, et al.
Published: (2025)
Adversarial Imitation Learning via Boosting
by: Chang, Jonathan D., et al.
Published: (2024)
by: Chang, Jonathan D., et al.
Published: (2024)
An Optimal Discriminator Weighted Imitation Perspective for Reinforcement Learning
by: Xu, Haoran, et al.
Published: (2025)
by: Xu, Haoran, et al.
Published: (2025)
Efficient Online RL Fine Tuning with Offline Pre-trained Policy Only
by: Xiao, Wei, et al.
Published: (2025)
by: Xiao, Wei, et al.
Published: (2025)
Boolean Satisfiability via Imitation Learning
by: Zhang, Zewei, et al.
Published: (2025)
by: Zhang, Zewei, et al.
Published: (2025)
Q-WSL: Optimizing Goal-Conditioned RL with Weighted Supervised Learning via Dynamic Programming
by: Lei, Xing, et al.
Published: (2024)
by: Lei, Xing, et al.
Published: (2024)
Diverse Policies Recovering via Pointwise Mutual Information Weighted Imitation Learning
by: Yang, Hanlin, et al.
Published: (2024)
by: Yang, Hanlin, et al.
Published: (2024)
Sample-efficient Adversarial Imitation Learning
by: Jung, Dahuin, et al.
Published: (2023)
by: Jung, Dahuin, et al.
Published: (2023)
Boundary-to-Region Supervision for Offline Safe Reinforcement Learning
by: Su, Huikang, et al.
Published: (2025)
by: Su, Huikang, et al.
Published: (2025)
Learning from Teaching Regularization: Generalizable Correlations Should be Easy to Imitate
by: Jin, Can, et al.
Published: (2024)
by: Jin, Can, et al.
Published: (2024)
On Learning Informative Trajectory Embeddings for Imitation, Classification and Regression
by: Ge, Zichang, et al.
Published: (2025)
by: Ge, Zichang, et al.
Published: (2025)
Unsupervised Motion Retargeting for Human-Robot Imitation
by: Annabi, Louis, et al.
Published: (2024)
by: Annabi, Louis, et al.
Published: (2024)
Diffusion-Reward Adversarial Imitation Learning
by: Lai, Chun-Mao, et al.
Published: (2024)
by: Lai, Chun-Mao, et al.
Published: (2024)
Deep Reinforcement Learning for Artificial Upwelling Energy Management
by: Zhang, Yiyuan, et al.
Published: (2023)
by: Zhang, Yiyuan, et al.
Published: (2023)
Tracking the Copyright of Large Vision-Language Models through Parameter Learning Adversarial Images
by: Wang, Yubo, et al.
Published: (2025)
by: Wang, Yubo, et al.
Published: (2025)
SENSOR: Imitate Third-Person Expert's Behaviors via Active Sensoring
by: Huang, Kaichen, et al.
Published: (2024)
by: Huang, Kaichen, et al.
Published: (2024)
Score-Based Diffusion Policy Compatible with Reinforcement Learning via Optimal Transport
by: Sun, Mingyang, et al.
Published: (2025)
by: Sun, Mingyang, et al.
Published: (2025)
DIDA: Denoised Imitation Learning based on Domain Adaptation
by: Huang, Kaichen, et al.
Published: (2024)
by: Huang, Kaichen, et al.
Published: (2024)
Adaptive Dual-Weighting Framework for Federated Learning via Out-of-Distribution Detection
by: Ling, Zhiwei, et al.
Published: (2026)
by: Ling, Zhiwei, et al.
Published: (2026)
Diversified Scaling Inference in Time Series Foundation Models
by: Hua, Ruijin, et al.
Published: (2026)
by: Hua, Ruijin, et al.
Published: (2026)
Reinformer: Max-Return Sequence Modeling for Offline RL
by: Zhuang, Zifeng, et al.
Published: (2024)
by: Zhuang, Zifeng, et al.
Published: (2024)
PAIL: Performance based Adversarial Imitation Learning Engine for Carbon Neutral Optimization
by: Ye, Yuyang, et al.
Published: (2024)
by: Ye, Yuyang, et al.
Published: (2024)
ScatterAD: Temporal-Topological Scattering Mechanism for Time Series Anomaly Detection
by: Yin, Tao, et al.
Published: (2025)
by: Yin, Tao, et al.
Published: (2025)
Imitation Learning via Focused Satisficing
by: Shah, Rushit N., et al.
Published: (2025)
by: Shah, Rushit N., et al.
Published: (2025)
Asymmetric On-Policy Distillation: Bridging Exploitation and Imitation at the Token Level
by: Jia, Nan, et al.
Published: (2026)
by: Jia, Nan, et al.
Published: (2026)
Disentangling Policy from Offline Task Representation Learning via Adversarial Data Augmentation
by: Jia, Chengxing, et al.
Published: (2024)
by: Jia, Chengxing, et al.
Published: (2024)
Denoising-based Contractive Imitation Learning
by: Shen, Macheng, et al.
Published: (2025)
by: Shen, Macheng, et al.
Published: (2025)
Adversarial Generative Flow Network for Solving Vehicle Routing Problems
by: Zhang, Ni, et al.
Published: (2025)
by: Zhang, Ni, et al.
Published: (2025)
Identifying Sensitive Weights via Post-quantization Integral
by: Hu, Yuezhou, et al.
Published: (2025)
by: Hu, Yuezhou, et al.
Published: (2025)
PAR-AdvGAN: Improving Adversarial Attack Capability with Progressive Auto-Regression AdvGAN
by: Zhang, Jiayu, et al.
Published: (2025)
by: Zhang, Jiayu, et al.
Published: (2025)
Beyond Independent Manipulation: Individual Fairness-aware Strategic Classification with Peer Imitation
by: Lv, Xinpeng, et al.
Published: (2026)
by: Lv, Xinpeng, et al.
Published: (2026)
COSEE: Consistency-Oriented Signal-Based Early Exiting via Calibrated Sample Weighting Mechanism
by: He, Jianing, et al.
Published: (2024)
by: He, Jianing, et al.
Published: (2024)
Reinforcement Learning via Implicit Imitation Guidance
by: Dong, Perry, et al.
Published: (2025)
by: Dong, Perry, et al.
Published: (2025)
Learning for Long-Horizon Planning via Neuro-Symbolic Abductive Imitation
by: Shao, Jie-Jing, et al.
Published: (2024)
by: Shao, Jie-Jing, et al.
Published: (2024)
Imitation Learning from Suboptimal Demonstrations via Meta-Learning An Action Ranker
by: Fan, Jiangdong, et al.
Published: (2024)
by: Fan, Jiangdong, et al.
Published: (2024)
Similar Items
-
Context-Former: Stitching via Latent Conditioned Sequence Modeling
by: Zhang, Ziqi, et al.
Published: (2024) -
A dynamical clipping approach with task feedback for Proximal Policy Optimization
by: Zhang, Ziqi, et al.
Published: (2023) -
Improving Offline-to-Online Reinforcement Learning with Q Conditioned State Entropy Exploration
by: Zhang, Ziqi, et al.
Published: (2023) -
Language Models as Efficient Reward Function Searchers for Custom-Environment Multi-Objective Reinforcement
by: Xie, Guanwen, et al.
Published: (2024) -
Self-Training with Dynamic Weighting for Robust Gradual Domain Adaptation
by: Wang, Zixi, et al.
Published: (2025)