Goal Inference from Open-Ended Dialog
Fuente:
arXiv
Saved in:
| Main Authors: | Ma, Rachel, Qu, Jingyi, Bobu, Andreea, Hadfield-Menell, Dylan |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Flexible Agent Alignment with Goal Inference from Open-Ended Dialog
by: Ma, Rachel, et al.
Published: (2025)
by: Ma, Rachel, et al.
Published: (2025)
Distributional Process Reward Models: Calibrated Prediction of Future Rewards via Conditional Optimal Transport
by: Ma, Rachel, et al.
Published: (2026)
by: Ma, Rachel, et al.
Published: (2026)
Robots That Know What to Ask: Recovering Misaligned Rewards through Targeted Explanations
by: Merker, Helena, et al.
Published: (2026)
by: Merker, Helena, et al.
Published: (2026)
Aligning Robot and Human Representations
by: Bobu, Andreea, et al.
Published: (2023)
by: Bobu, Andreea, et al.
Published: (2023)
Distributional Preference Learning: Understanding and Accounting for Hidden Context in RLHF
by: Siththaranjan, Anand, et al.
Published: (2023)
by: Siththaranjan, Anand, et al.
Published: (2023)
Alignment Tampering: How Reinforcement Learning from Human Feedback Is Exploited to Optimize Misaligned Biases
by: Hahm, Dongyoon, et al.
Published: (2026)
by: Hahm, Dongyoon, et al.
Published: (2026)
Enabling Adaptive Agent Training in Open-Ended Simulators by Targeting Diversity
by: Costales, Robby, et al.
Published: (2024)
by: Costales, Robby, et al.
Published: (2024)
Preference-Conditioned Language-Guided Abstraction
by: Peng, Andi, et al.
Published: (2024)
by: Peng, Andi, et al.
Published: (2024)
The Autonomy-Alignment Problem in Open-Ended Learning Robots: Formalising the Purpose Framework
by: Baldassarre, Gianluca, et al.
Published: (2024)
by: Baldassarre, Gianluca, et al.
Published: (2024)
FOUNDER: Grounding Foundation Models in World Models for Open-Ended Embodied Decision Making
by: Wang, Yucen, et al.
Published: (2025)
by: Wang, Yucen, et al.
Published: (2025)
Hi Robot: Open-Ended Instruction Following with Hierarchical Vision-Language-Action Models
by: Shi, Lucy Xiaoyang, et al.
Published: (2025)
by: Shi, Lucy Xiaoyang, et al.
Published: (2025)
Defending Against Unforeseen Failure Modes with Latent Adversarial Training
by: Casper, Stephen, et al.
Published: (2024)
by: Casper, Stephen, et al.
Published: (2024)
Open-Ended Goal Inference through Actions and Language for Human-Robot Collaboration
by: Ghose, Debasmita, et al.
Published: (2025)
by: Ghose, Debasmita, et al.
Published: (2025)
Masked IRL: LLM-Guided Reward Disambiguation from Demonstrations and Language
by: Hwang, Minyoung, et al.
Published: (2025)
by: Hwang, Minyoung, et al.
Published: (2025)
Robometer: Scaling General-Purpose Robotic Reward Models via Trajectory Comparisons
by: Liang, Anthony, et al.
Published: (2026)
by: Liang, Anthony, et al.
Published: (2026)
QuickLAP: Quick Language-Action Preference Learning for Semi-Autonomous Agents
by: Nader, Jordan Abi, et al.
Published: (2025)
by: Nader, Jordan Abi, et al.
Published: (2025)
Reconciling Spatial and Temporal Abstractions for Goal Representation
by: Zadem, Mehdi, et al.
Published: (2024)
by: Zadem, Mehdi, et al.
Published: (2024)
Learning World Models for Unconstrained Goal Navigation
by: Duan, Yuanlin, et al.
Published: (2024)
by: Duan, Yuanlin, et al.
Published: (2024)
Goal-Conditioned Data Augmentation for Offline Reinforcement Learning
by: Huang, Xingshuai, et al.
Published: (2024)
by: Huang, Xingshuai, et al.
Published: (2024)
GRAIL: Goal Recognition Alignment through Imitation Learning
by: Elhadad, Osher, et al.
Published: (2026)
by: Elhadad, Osher, et al.
Published: (2026)
SMORE: Score Models for Offline Goal-Conditioned Reinforcement Learning
by: Sikchi, Harshit, et al.
Published: (2023)
by: Sikchi, Harshit, et al.
Published: (2023)
HIQL: Offline Goal-Conditioned RL with Latent States as Actions
by: Park, Seohong, et al.
Published: (2023)
by: Park, Seohong, et al.
Published: (2023)
Learning to Open and Traverse Doors with a Legged Manipulator
by: Zhang, Mike, et al.
Published: (2024)
by: Zhang, Mike, et al.
Published: (2024)
STEVE-Audio: Expanding the Goal Conditioning Modalities of Embodied Agents in Minecraft
by: Lenzen, Nicholas, et al.
Published: (2024)
by: Lenzen, Nicholas, et al.
Published: (2024)
Exploring the Edges of Latent State Clusters for Goal-Conditioned Reinforcement Learning
by: Duan, Yuanlin, et al.
Published: (2024)
by: Duan, Yuanlin, et al.
Published: (2024)
Dynamic Neural Curiosity Enhances Learning Flexibility for Autonomous Goal Discovery
by: Houbre, Quentin, et al.
Published: (2024)
by: Houbre, Quentin, et al.
Published: (2024)
Visualizing Critic Match Loss Landscapes for Interpretation of Online Reinforcement Learning Control Algorithms
by: Liu, Jingyi, et al.
Published: (2026)
by: Liu, Jingyi, et al.
Published: (2026)
Adapting Critic Match Loss Landscape Visualization to Off-policy Reinforcement Learning
by: Liu, Jingyi, et al.
Published: (2026)
by: Liu, Jingyi, et al.
Published: (2026)
A Loss Landscape Visualization Framework for Interpreting Reinforcement Learning: An ADHDP Case Study
by: Liu, Jingyi, et al.
Published: (2026)
by: Liu, Jingyi, et al.
Published: (2026)
Slot-MPC: Goal-Conditioned Model Predictive Control with Object-Centric Representations
by: Spieler, Jonathan, et al.
Published: (2026)
by: Spieler, Jonathan, et al.
Published: (2026)
Data-Efficient Hierarchical Goal-Conditioned Reinforcement Learning via Normalizing Flows
by: Garg, Shaswat, et al.
Published: (2026)
by: Garg, Shaswat, et al.
Published: (2026)
Real Evaluations Tractability using Continuous Goal-Directed Actions in Smart City Applications
by: Fernandez-Fernandez, Raul, et al.
Published: (2024)
by: Fernandez-Fernandez, Raul, et al.
Published: (2024)
Efficient Virtuoso: A Latent Diffusion Transformer Model for Goal-Conditioned Trajectory Planning
by: Guillen-Perez, Antonio
Published: (2025)
by: Guillen-Perez, Antonio
Published: (2025)
Unsupervised Learning of Efficient Exploration: Pre-training Adaptive Policies via Self-Imposed Goals
by: Pappalardo, Octavio
Published: (2026)
by: Pappalardo, Octavio
Published: (2026)
A Goal-Oriented Reinforcement Learning-Based Path Planning Algorithm for Modular Self-Reconfigurable Satellites
by: Liu, Bofei, et al.
Published: (2025)
by: Liu, Bofei, et al.
Published: (2025)
A Data-Based Architecture for Flight Test without Test Points
by: Harp, D. Isaiah, et al.
Published: (2025)
by: Harp, D. Isaiah, et al.
Published: (2025)
Rule-based High-Level Coaching for Goal-Conditioned Reinforcement Learning in Search-and-Rescue UAV Missions Under Limited-Simulation Training
by: Ramezani, Mahya, et al.
Published: (2026)
by: Ramezani, Mahya, et al.
Published: (2026)
CHARMS: A Cognitive Hierarchical Agent for Reasoning and Motion Stylization in Autonomous Driving
by: Wang, Jingyi, et al.
Published: (2025)
by: Wang, Jingyi, et al.
Published: (2025)
EvtSlowTV -- A Large and Diverse Dataset for Event-Based Depth Estimation
by: Macaulay, Sadiq Layi, et al.
Published: (2025)
by: Macaulay, Sadiq Layi, et al.
Published: (2025)
Learning How Hard to Think: Input-Adaptive Allocation of LM Computation
by: Damani, Mehul, et al.
Published: (2024)
by: Damani, Mehul, et al.
Published: (2024)
Similar Items
-
Flexible Agent Alignment with Goal Inference from Open-Ended Dialog
by: Ma, Rachel, et al.
Published: (2025) -
Distributional Process Reward Models: Calibrated Prediction of Future Rewards via Conditional Optimal Transport
by: Ma, Rachel, et al.
Published: (2026) -
Robots That Know What to Ask: Recovering Misaligned Rewards through Targeted Explanations
by: Merker, Helena, et al.
Published: (2026) -
Aligning Robot and Human Representations
by: Bobu, Andreea, et al.
Published: (2023) -
Distributional Preference Learning: Understanding and Accounting for Hidden Context in RLHF
by: Siththaranjan, Anand, et al.
Published: (2023)