Saved in:
| Main Authors: | Hao, Jianye, Yuan, Yifu, Wang, Cong, Wang, Zhen |
|---|---|
| Format: | Preprint |
| Published: |
2021
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2112.02817 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MODULI: Unlocking Preference Generalization via Diffusion Models for Offline Multi-Objective Reinforcement Learning
by: Yuan, Yifu, et al.
Published: (2024)
by: Yuan, Yifu, et al.
Published: (2024)
MENTOR: Guiding Hierarchical Reinforcement Learning with Human Feedback and Dynamic Distance Constraint
by: Zhou, Xinglin, et al.
Published: (2024)
by: Zhou, Xinglin, et al.
Published: (2024)
AhaRobot: A Low-Cost Open-Source Bimanual Mobile Manipulator for Embodied AI
by: Cui, Haiqin, et al.
Published: (2025)
by: Cui, Haiqin, et al.
Published: (2025)
Enhancing Robotic Manipulation with AI Feedback from Multimodal Large Language Models
by: Liu, Jinyi, et al.
Published: (2024)
by: Liu, Jinyi, et al.
Published: (2024)
CleanDiffuser: An Easy-to-use Modularized Library for Diffusion Models in Decision Making
by: Dong, Zibin, et al.
Published: (2024)
by: Dong, Zibin, et al.
Published: (2024)
Learnable Behavior Control: Breaking Atari Human World Records via Sample-Efficient Behavior Selection
by: Fan, Jiajun, et al.
Published: (2023)
by: Fan, Jiajun, et al.
Published: (2023)
SheetAgent: Towards A Generalist Agent for Spreadsheet Reasoning and Manipulation via Large Language Models
by: Chen, Yibin, et al.
Published: (2024)
by: Chen, Yibin, et al.
Published: (2024)
ED-VAE: Entropy Decomposition of ELBO in Variational Autoencoders
by: Lygerakis, Fotios, et al.
Published: (2024)
by: Lygerakis, Fotios, et al.
Published: (2024)
Pessimistic Value Iteration for Multi-Task Data Sharing in Offline Reinforcement Learning
by: Bai, Chenjia, et al.
Published: (2024)
by: Bai, Chenjia, et al.
Published: (2024)
TD-MPC2: Scalable, Robust World Models for Continuous Control
by: Hansen, Nicklas, et al.
Published: (2023)
by: Hansen, Nicklas, et al.
Published: (2023)
DecoKAN: Interpretable Decomposition for Forecasting Cryptocurrency Market Dynamics
by: Gao, Yuan, et al.
Published: (2025)
by: Gao, Yuan, et al.
Published: (2025)
Embodied-R1: Reinforced Embodied Reasoning for General Robotic Manipulation
by: Yuan, Yifu, et al.
Published: (2025)
by: Yuan, Yifu, et al.
Published: (2025)
From Seeing to Doing: Bridging Reasoning and Decision for Robotic Manipulation
by: Yuan, Yifu, et al.
Published: (2025)
by: Yuan, Yifu, et al.
Published: (2025)
Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning
by: Wu, Qingyuan, et al.
Published: (2025)
by: Wu, Qingyuan, et al.
Published: (2025)
ED-Filter: Dynamic Feature Filtering for Eating Disorder Classification
by: Naseriparsa, Mehdi, et al.
Published: (2025)
by: Naseriparsa, Mehdi, et al.
Published: (2025)
Deterministic Decomposition of Stochastic Generative Dynamics
by: Song, Xingyu, et al.
Published: (2026)
by: Song, Xingyu, et al.
Published: (2026)
Uni-RLHF: Universal Platform and Benchmark Suite for Reinforcement Learning with Diverse Human Feedback
by: Yuan, Yifu, et al.
Published: (2024)
by: Yuan, Yifu, et al.
Published: (2024)
VDFD: Multi-Agent Value Decomposition Framework with Disentangled World Model
by: Wang, Zhizun, et al.
Published: (2023)
by: Wang, Zhizun, et al.
Published: (2023)
Reflex: Reinforcement Learning with Reflection Symmetry Exploitation in State-Based Continuous Control
by: Zhen, Shuai, et al.
Published: (2026)
by: Zhen, Shuai, et al.
Published: (2026)
Wavelet Fourier Diffuser: Frequency-Aware Diffusion Model for Reinforcement Learning
by: Luo, Yifu, et al.
Published: (2025)
by: Luo, Yifu, et al.
Published: (2025)
Self-Controlled Dynamic Expansion Model for Continual Learning
by: Wu, Runqing, et al.
Published: (2025)
by: Wu, Runqing, et al.
Published: (2025)
CLeAN: Continual Learning Adaptive Normalization in Dynamic Environments
by: Marasco, Isabella, et al.
Published: (2026)
by: Marasco, Isabella, et al.
Published: (2026)
AndroidWorld: A Dynamic Benchmarking Environment for Autonomous Agents
by: Rawles, Christopher, et al.
Published: (2024)
by: Rawles, Christopher, et al.
Published: (2024)
GraSS: Combining Graph Neural Networks with Expert Knowledge for SAT Solver Selection
by: Zhang, Zhanguang, et al.
Published: (2024)
by: Zhang, Zhanguang, et al.
Published: (2024)
WIMLE: Uncertainty-Aware World Models with IMLE for Sample-Efficient Continuous Control
by: Aghabozorgi, Mehran, et al.
Published: (2026)
by: Aghabozorgi, Mehran, et al.
Published: (2026)
DistRL: An Asynchronous Distributed Reinforcement Learning Framework for On-Device Control Agents
by: Wang, Taiyi, et al.
Published: (2024)
by: Wang, Taiyi, et al.
Published: (2024)
Ratio-Variance Regularized Policy Optimization for Efficient LLM Fine-tuning
by: Luo, Yu, et al.
Published: (2026)
by: Luo, Yu, et al.
Published: (2026)
SeaDAG: Semi-autoregressive Diffusion for Conditional Directed Acyclic Graph Generation
by: Zhou, Xinyi, et al.
Published: (2024)
by: Zhou, Xinyi, et al.
Published: (2024)
Continual Learning for Adaptable Car-Following in Dynamic Traffic Environments
by: Chen, Xianda, et al.
Published: (2024)
by: Chen, Xianda, et al.
Published: (2024)
Skewed Memorization in Large Language Models: Quantification and Decomposition
by: Li, Hao, et al.
Published: (2025)
by: Li, Hao, et al.
Published: (2025)
Self-Improving World Modelling with Latent Actions
by: Qiu, Yifu, et al.
Published: (2026)
by: Qiu, Yifu, et al.
Published: (2026)
Benchmarking World-Model Learning with Environment-Level Queries
by: Warrier, Archana, et al.
Published: (2025)
by: Warrier, Archana, et al.
Published: (2025)
FedMobile: Enabling Knowledge Contribution-aware Multi-modal Federated Learning with Incomplete Modalities
by: Liu, Yi, et al.
Published: (2025)
by: Liu, Yi, et al.
Published: (2025)
Agent World Model: Infinity Synthetic Environments for Agentic Reinforcement Learning
by: Wang, Zhaoyang, et al.
Published: (2026)
by: Wang, Zhaoyang, et al.
Published: (2026)
Adaptive Overclocking: Dynamic Control of Thinking Path Length via Real-Time Reasoning Signals
by: Jiang, Shuhao, et al.
Published: (2025)
by: Jiang, Shuhao, et al.
Published: (2025)
Scalable In-Context Q-Learning
by: Liu, Jinmei, et al.
Published: (2025)
by: Liu, Jinmei, et al.
Published: (2025)
Policy and World Modeling Co-Training for Language Agents
by: Lu, Ning, et al.
Published: (2026)
by: Lu, Ning, et al.
Published: (2026)
Generative Models in Decision Making: A Survey
by: Shao, Xinyu, et al.
Published: (2025)
by: Shao, Xinyu, et al.
Published: (2025)
Spatiotemporal Forecasting as Planning: A Model-Based Reinforcement Learning Approach with Generative World Models
by: Wu, Hao, et al.
Published: (2025)
by: Wu, Hao, et al.
Published: (2025)
Condensation-Concatenation Framework for Dynamic Graph Continual Learning
by: Yan, Tingxu, et al.
Published: (2025)
by: Yan, Tingxu, et al.
Published: (2025)
Similar Items
-
MODULI: Unlocking Preference Generalization via Diffusion Models for Offline Multi-Objective Reinforcement Learning
by: Yuan, Yifu, et al.
Published: (2024) -
MENTOR: Guiding Hierarchical Reinforcement Learning with Human Feedback and Dynamic Distance Constraint
by: Zhou, Xinglin, et al.
Published: (2024) -
AhaRobot: A Low-Cost Open-Source Bimanual Mobile Manipulator for Embodied AI
by: Cui, Haiqin, et al.
Published: (2025) -
Enhancing Robotic Manipulation with AI Feedback from Multimodal Large Language Models
by: Liu, Jinyi, et al.
Published: (2024) -
CleanDiffuser: An Easy-to-use Modularized Library for Diffusion Models in Decision Making
by: Dong, Zibin, et al.
Published: (2024)