Rewarding DINO: Predicting Dense Rewards with Vision Foundation Models
Fuente:
arXiv
Saved in:
| Main Authors: | Krack, Pierre, Jülg, Tobias, Burgard, Wolfram, Walter, Florian |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Refined Policy Distillation: From VLA Generalists to RL Experts
by: Jülg, Tobias, et al.
Published: (2025)
by: Jülg, Tobias, et al.
Published: (2025)
Robot Control Stack: A Lean Ecosystem for Robot Learning at Scale
by: Jülg, Tobias, et al.
Published: (2025)
by: Jülg, Tobias, et al.
Published: (2025)
LLM-Pack: Intuitive Grocery Handling for Logistics Applications
by: Blei, Yannik, et al.
Published: (2025)
by: Blei, Yannik, et al.
Published: (2025)
FlowTouch: View-Invariant Visuo-Tactile Prediction
by: Bien, Seongjin, et al.
Published: (2026)
by: Bien, Seongjin, et al.
Published: (2026)
Augmented Reality for RObots (ARRO): Pointing Visuomotor Policies Towards Visual Robustness
by: Mirjalili, Reihaneh, et al.
Published: (2025)
by: Mirjalili, Reihaneh, et al.
Published: (2025)
VLAgents: A Policy Server for Efficient VLA Inference
by: Jülg, Tobias, et al.
Published: (2026)
by: Jülg, Tobias, et al.
Published: (2026)
Agent-Agnostic Centralized Training for Decentralized Multi-Agent Cooperative Driving
by: Yan, Shengchao, et al.
Published: (2024)
by: Yan, Shengchao, et al.
Published: (2024)
TimeRewarder: Learning Dense Reward from Passive Videos via Frame-wise Temporal Distance
by: Liu, Yuyang, et al.
Published: (2025)
by: Liu, Yuyang, et al.
Published: (2025)
DrS: Learning Reusable Dense Rewards for Multi-Stage Tasks
by: Mu, Tongzhou, et al.
Published: (2024)
by: Mu, Tongzhou, et al.
Published: (2024)
Policy Learning from Large Vision-Language Model Feedback without Reward Modeling
by: Luu, Tung M., et al.
Published: (2025)
by: Luu, Tung M., et al.
Published: (2025)
ELEMENTAL: Interactive Learning from Demonstrations and Vision-Language Models for Reward Design in Robotics
by: Chen, Letian, et al.
Published: (2024)
by: Chen, Letian, et al.
Published: (2024)
Generalizable Dense Reward for Long-Horizon Robotic Tasks
by: Yong, Silong, et al.
Published: (2026)
by: Yong, Silong, et al.
Published: (2026)
The Dark Side of Rich Rewards: Understanding and Mitigating Noise in VLM Rewards
by: Huang, Sukai, et al.
Published: (2024)
by: Huang, Sukai, et al.
Published: (2024)
Reward Prediction Error Prioritisation in Experience Replay: The RPE-PER Method
by: Yamani, Hoda, et al.
Published: (2025)
by: Yamani, Hoda, et al.
Published: (2025)
From Demonstrations to Rewards: Test-Time Prompt Optimization for VLM Reward Models
by: Gumbsch, Christian, et al.
Published: (2026)
by: Gumbsch, Christian, et al.
Published: (2026)
Subwords as Skills: Tokenization for Sparse-Reward Reinforcement Learning
by: Yunis, David, et al.
Published: (2023)
by: Yunis, David, et al.
Published: (2023)
BYE: Build Your Encoder with One Sequence of Exploration Data for Long-Term Dynamic Scene Understanding
by: Huang, Chenguang, et al.
Published: (2024)
by: Huang, Chenguang, et al.
Published: (2024)
RobotKeyframing: Learning Locomotion with High-Level Objectives via Mixture of Dense and Sparse Rewards
by: Zargarbashi, Fatemeh, et al.
Published: (2024)
by: Zargarbashi, Fatemeh, et al.
Published: (2024)
Constraints as Rewards: Reinforcement Learning for Robots without Reward Functions
by: Ishihara, Yu, et al.
Published: (2025)
by: Ishihara, Yu, et al.
Published: (2025)
Uncertainty-aware Reward Design Process
by: Yang, Yang, et al.
Published: (2025)
by: Yang, Yang, et al.
Published: (2025)
Reward Machine Inference for Robotic Manipulation
by: Baert, Mattijs, et al.
Published: (2024)
by: Baert, Mattijs, et al.
Published: (2024)
ReLAM: Learning Anticipation Model for Rewarding Visual Robotic Manipulation
by: Tang, Nan, et al.
Published: (2025)
by: Tang, Nan, et al.
Published: (2025)
Text2Reward: Reward Shaping with Language Models for Reinforcement Learning
by: Xie, Tianbao, et al.
Published: (2023)
by: Xie, Tianbao, et al.
Published: (2023)
Curriculum Reinforcement Learning for Complex Reward Functions
by: Freitag, Kilian, et al.
Published: (2024)
by: Freitag, Kilian, et al.
Published: (2024)
On-Robot Reinforcement Learning with Goal-Contrastive Rewards
by: Biza, Ondrej, et al.
Published: (2024)
by: Biza, Ondrej, et al.
Published: (2024)
ORSO: Accelerating Reward Design via Online Reward Selection and Policy Optimization
by: Zhang, Chen Bo Calvin, et al.
Published: (2024)
by: Zhang, Chen Bo Calvin, et al.
Published: (2024)
Revisiting Sparse Rewards for Goal-Reaching Reinforcement Learning
by: Vasan, Gautham, et al.
Published: (2024)
by: Vasan, Gautham, et al.
Published: (2024)
Reward Redistribution via Gaussian Process Likelihood Estimation
by: Xiao, Minheng, et al.
Published: (2025)
by: Xiao, Minheng, et al.
Published: (2025)
Residual Reward Models for Preference-based Reinforcement Learning
by: Cao, Chenyang, et al.
Published: (2025)
by: Cao, Chenyang, et al.
Published: (2025)
Generalization in Deep Reinforcement Learning for Robotic Navigation by Reward Shaping
by: Miranda, Victor R. F., et al.
Published: (2022)
by: Miranda, Victor R. F., et al.
Published: (2022)
Diffusion Reward: Learning Rewards via Conditional Video Diffusion
by: Huang, Tao, et al.
Published: (2023)
by: Huang, Tao, et al.
Published: (2023)
VLM-Vac: Enhancing Smart Vacuums through VLM Knowledge Distillation and Language-Guided Experience Replay
by: Mirjalili, Reihaneh, et al.
Published: (2024)
by: Mirjalili, Reihaneh, et al.
Published: (2024)
DiWA: Diffusion Policy Adaptation with World Models
by: Chandra, Akshay L, et al.
Published: (2025)
by: Chandra, Akshay L, et al.
Published: (2025)
Diffusion-Reward Adversarial Imitation Learning
by: Lai, Chun-Mao, et al.
Published: (2024)
by: Lai, Chun-Mao, et al.
Published: (2024)
ProcVLM: Learning Procedure-Grounded Progress Rewards for Robotic Manipulation
by: Feng, Youhe, et al.
Published: (2026)
by: Feng, Youhe, et al.
Published: (2026)
Momentum Based Reward Design for Low Emission Traffic Signal Control
by: Mundane, Chinmay, et al.
Published: (2026)
by: Mundane, Chinmay, et al.
Published: (2026)
Adaptive Teaching in Heterogeneous Agents: Balancing Surprise in Sparse Reward Scenarios
by: Clark, Emma, et al.
Published: (2024)
by: Clark, Emma, et al.
Published: (2024)
TopoNav: Topological Navigation for Efficient Exploration in Sparse Reward Environments
by: Hossain, Jumman, et al.
Published: (2024)
by: Hossain, Jumman, et al.
Published: (2024)
Bridging the Human to Robot Dexterity Gap through Object-Oriented Rewards
by: Guzey, Irmak, et al.
Published: (2024)
by: Guzey, Irmak, et al.
Published: (2024)
LUMOS: Language-Conditioned Imitation Learning with World Models
by: Nematollahi, Iman, et al.
Published: (2025)
by: Nematollahi, Iman, et al.
Published: (2025)
Similar Items
-
Refined Policy Distillation: From VLA Generalists to RL Experts
by: Jülg, Tobias, et al.
Published: (2025) -
Robot Control Stack: A Lean Ecosystem for Robot Learning at Scale
by: Jülg, Tobias, et al.
Published: (2025) -
LLM-Pack: Intuitive Grocery Handling for Logistics Applications
by: Blei, Yannik, et al.
Published: (2025) -
FlowTouch: View-Invariant Visuo-Tactile Prediction
by: Bien, Seongjin, et al.
Published: (2026) -
Augmented Reality for RObots (ARRO): Pointing Visuomotor Policies Towards Visual Robustness
by: Mirjalili, Reihaneh, et al.
Published: (2025)