Goal-Driven Reward by Video Diffusion Models for Reinforcement Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Qi, Wu, Mian, Zhang, Yuyang, Yuan, Mingqi, Zhang, Wenyao, You, Haoxiang, Wang, Yunbo, Jin, Xin, Yang, Xiaokang, Zeng, Wenjun |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Deep Reinforcement Learning with Hybrid Intrinsic Reward Model
von: Yuan, Mingqi, et al.
Veröffentlicht: (2025)
von: Yuan, Mingqi, et al.
Veröffentlicht: (2025)
Making Offline RL Online: Collaborative World Models for Offline Visual Reinforcement Learning
von: Wang, Qi, et al.
Veröffentlicht: (2023)
von: Wang, Qi, et al.
Veröffentlicht: (2023)
Disentangled World Models: Learning to Transfer Semantic Knowledge from Distracting Videos for Reinforcement Learning
von: Wang, Qi, et al.
Veröffentlicht: (2025)
von: Wang, Qi, et al.
Veröffentlicht: (2025)
Open-World Reinforcement Learning over Long Short-Term Imagination
von: Li, Jiajian, et al.
Veröffentlicht: (2024)
von: Li, Jiajian, et al.
Veröffentlicht: (2024)
Plasticine: Accelerating Research in Plasticity-Motivated Deep Reinforcement Learning
von: Yuan, Mingqi, et al.
Veröffentlicht: (2025)
von: Yuan, Mingqi, et al.
Veröffentlicht: (2025)
Adaptive Data Exploitation in Deep Reinforcement Learning
von: Yuan, Mingqi, et al.
Veröffentlicht: (2025)
von: Yuan, Mingqi, et al.
Veröffentlicht: (2025)
ULTHO: Ultra-Lightweight yet Efficient Hyperparameter Optimization in Deep Reinforcement Learning
von: Yuan, Mingqi, et al.
Veröffentlicht: (2025)
von: Yuan, Mingqi, et al.
Veröffentlicht: (2025)
RLLTE: Long-Term Evolution Project of Reinforcement Learning
von: Yuan, Mingqi, et al.
Veröffentlicht: (2023)
von: Yuan, Mingqi, et al.
Veröffentlicht: (2023)
Disentangled Robot Learning via Separate Forward and Inverse Dynamics Pretraining
von: Zhang, Wenyao, et al.
Veröffentlicht: (2026)
von: Zhang, Wenyao, et al.
Veröffentlicht: (2026)
Video-Enhanced Offline Reinforcement Learning: A Model-Based Approach
von: Pan, Minting, et al.
Veröffentlicht: (2025)
von: Pan, Minting, et al.
Veröffentlicht: (2025)
ReWorld: Multi-Dimensional Reward Modeling for Embodied World Models
von: Peng, Baorui, et al.
Veröffentlicht: (2026)
von: Peng, Baorui, et al.
Veröffentlicht: (2026)
RLeXplore: Accelerating Research in Intrinsically-Motivated Reinforcement Learning
von: Yuan, Mingqi, et al.
Veröffentlicht: (2024)
von: Yuan, Mingqi, et al.
Veröffentlicht: (2024)
Model-Based Reinforcement Learning with Multi-Task Offline Pretraining
von: Pan, Minting, et al.
Veröffentlicht: (2023)
von: Pan, Minting, et al.
Veröffentlicht: (2023)
Latent Intuitive Physics: Learning to Transfer Hidden Physics from A 3D Video
von: Zhu, Xiangming, et al.
Veröffentlicht: (2024)
von: Zhu, Xiangming, et al.
Veröffentlicht: (2024)
Improving Masked Autoencoders by Learning Where to Mask
von: Chen, Haijian, et al.
Veröffentlicht: (2023)
von: Chen, Haijian, et al.
Veröffentlicht: (2023)
MetaGS: A Meta-Learned Gaussian-Phong Model for Out-of-Distribution 3D Scene Relighting
von: He, Yumeng, et al.
Veröffentlicht: (2024)
von: He, Yumeng, et al.
Veröffentlicht: (2024)
Your Offline Policy is Not Trustworthy: Bilevel Reinforcement Learning for Sequential Portfolio Optimization
von: Yuan, Haochen, et al.
Veröffentlicht: (2025)
von: Yuan, Haochen, et al.
Veröffentlicht: (2025)
DimINO: Dimension-Informed Neural Operator Learning
von: Song, Yichen, et al.
Veröffentlicht: (2024)
von: Song, Yichen, et al.
Veröffentlicht: (2024)
TAMTRL: Teacher-Aligned Reward Reshaping for Multi-Turn Reinforcement Learning in Long-Context Compression
von: Wang, Li, et al.
Veröffentlicht: (2026)
von: Wang, Li, et al.
Veröffentlicht: (2026)
Closed-Loop Unsupervised Representation Disentanglement with $β$-VAE Distillation and Diffusion Probabilistic Feedback
von: Jin, Xin, et al.
Veröffentlicht: (2024)
von: Jin, Xin, et al.
Veröffentlicht: (2024)
Interpretable Single-View 3D Gaussian Splatting using Unsupervised Hierarchical Disentangled Representation Learning
von: Zhang, Yuyang, et al.
Veröffentlicht: (2025)
von: Zhang, Yuyang, et al.
Veröffentlicht: (2025)
Diffusion Classifier-Driven Reward for Offline Preference-based Reinforcement Learning
von: Pang, Teng, et al.
Veröffentlicht: (2025)
von: Pang, Teng, et al.
Veröffentlicht: (2025)
Continual Visual Reinforcement Learning with A Life-Long World Model
von: Pan, Minting, et al.
Veröffentlicht: (2023)
von: Pan, Minting, et al.
Veröffentlicht: (2023)
Unicorn: Scaling High-Dimensional Time Series Forecasting via Universal Correlation Modeling
von: Yuan, Haochen, et al.
Veröffentlicht: (2026)
von: Yuan, Haochen, et al.
Veröffentlicht: (2026)
Hierarchical Temporal Context Learning for Camera-based Semantic Scene Completion
von: Li, Bohan, et al.
Veröffentlicht: (2024)
von: Li, Bohan, et al.
Veröffentlicht: (2024)
DynaVol: Unsupervised Learning for Dynamic Scenes through Object-Centric Voxelization
von: Zhao, Yanpeng, et al.
Veröffentlicht: (2023)
von: Zhao, Yanpeng, et al.
Veröffentlicht: (2023)
PvP: Data-Efficient Humanoid Robot Learning with Proprioceptive-Privileged Contrastive Representations
von: Yuan, Mingqi, et al.
Veröffentlicht: (2025)
von: Yuan, Mingqi, et al.
Veröffentlicht: (2025)
Revisiting Sparse Rewards for Goal-Reaching Reinforcement Learning
von: Vasan, Gautham, et al.
Veröffentlicht: (2024)
von: Vasan, Gautham, et al.
Veröffentlicht: (2024)
ReAugment: Model Zoo-Guided RL for Few-Shot Time Series Augmentation and Forecasting
von: Yuan, Haochen, et al.
Veröffentlicht: (2024)
von: Yuan, Haochen, et al.
Veröffentlicht: (2024)
Scene Graph Disentanglement and Composition for Generalizable Complex Image Generation
von: Wang, Yunnan, et al.
Veröffentlicht: (2024)
von: Wang, Yunnan, et al.
Veröffentlicht: (2024)
On-Robot Reinforcement Learning with Goal-Contrastive Rewards
von: Biza, Ondrej, et al.
Veröffentlicht: (2024)
von: Biza, Ondrej, et al.
Veröffentlicht: (2024)
LASER: Learning Active Sensing for Continuum Field Reconstruction
von: Deng, Huayu, et al.
Veröffentlicht: (2026)
von: Deng, Huayu, et al.
Veröffentlicht: (2026)
EvoMesh: Adaptive Physical Simulation with Hierarchical Graph Evolutions
von: Deng, Huayu, et al.
Veröffentlicht: (2024)
von: Deng, Huayu, et al.
Veröffentlicht: (2024)
AIM: Intent-Aware Unified world action Modeling with Spatial Value Maps
von: Fan, Liaoyuan, et al.
Veröffentlicht: (2026)
von: Fan, Liaoyuan, et al.
Veröffentlicht: (2026)
OrbiSim: World Models as Differentiable Physics Engines for Embodied Intelligence
von: Li, Jiajian, et al.
Veröffentlicht: (2026)
von: Li, Jiajian, et al.
Veröffentlicht: (2026)
From Sparse to Dense: Toddler-inspired Reward Transition in Goal-Oriented Reinforcement Learning
von: Park, Junseok, et al.
Veröffentlicht: (2025)
von: Park, Junseok, et al.
Veröffentlicht: (2025)
Hybrid-grained Feature Aggregation with Coarse-to-fine Language Guidance for Self-supervised Monocular Depth Estimation
von: Zhang, Wenyao, et al.
Veröffentlicht: (2025)
von: Zhang, Wenyao, et al.
Veröffentlicht: (2025)
Bridging Visual Representation and Reinforcement Learning from Verifiable Rewards in Large Vision-Language Models
von: Han, Yuhang, et al.
Veröffentlicht: (2026)
von: Han, Yuhang, et al.
Veröffentlicht: (2026)
Goal Conditioned Reinforcement Learning for Photo Finishing Tuning
von: Wu, Jiarui, et al.
Veröffentlicht: (2025)
von: Wu, Jiarui, et al.
Veröffentlicht: (2025)
MVR: Multi-view Video Reward Shaping for Reinforcement Learning
von: Luo, Lirui, et al.
Veröffentlicht: (2026)
von: Luo, Lirui, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Deep Reinforcement Learning with Hybrid Intrinsic Reward Model
von: Yuan, Mingqi, et al.
Veröffentlicht: (2025) -
Making Offline RL Online: Collaborative World Models for Offline Visual Reinforcement Learning
von: Wang, Qi, et al.
Veröffentlicht: (2023) -
Disentangled World Models: Learning to Transfer Semantic Knowledge from Distracting Videos for Reinforcement Learning
von: Wang, Qi, et al.
Veröffentlicht: (2025) -
Open-World Reinforcement Learning over Long Short-Term Imagination
von: Li, Jiajian, et al.
Veröffentlicht: (2024) -
Plasticine: Accelerating Research in Plasticity-Motivated Deep Reinforcement Learning
von: Yuan, Mingqi, et al.
Veröffentlicht: (2025)