Confidence-Controlled Exploration: Efficient Sparse-Reward Policy Learning for Robot Navigation
Fuente:
arXiv
Salvato in:
| Autori principali: | Patel, Bhrij, Weerakoon, Kasun, Suttle, Wesley A., Koppel, Alec, Sadler, Brian M., Zhou, Tianyi, Bedi, Amrit Singh, Manocha, Dinesh |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2023
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Towards Global Optimality for Practical Average Reward Reinforcement Learning without Mixing Time Oracles
di: Patel, Bhrij, et al.
Pubblicazione: (2024)
di: Patel, Bhrij, et al.
Pubblicazione: (2024)
Multi-LLM QA with Embodied Exploration
di: Patel, Bhrij, et al.
Pubblicazione: (2024)
di: Patel, Bhrij, et al.
Pubblicazione: (2024)
Personalized Embodied Navigation for Portable Object Finding
di: Dorbala, Vishnu Sashank, et al.
Pubblicazione: (2024)
di: Dorbala, Vishnu Sashank, et al.
Pubblicazione: (2024)
Code Comprehension then Auditing for Unsupervised LLM Evaluation
di: Patel, Bhrij, et al.
Pubblicazione: (2024)
di: Patel, Bhrij, et al.
Pubblicazione: (2024)
ProNav: Proprioceptive Traversability Estimation for Legged Robot Navigation in Outdoor Environments
di: Elnoor, Mohamed, et al.
Pubblicazione: (2023)
di: Elnoor, Mohamed, et al.
Pubblicazione: (2023)
PARL: A Unified Framework for Policy Alignment in Reinforcement Learning from Human Feedback
di: Chakraborty, Souradip, et al.
Pubblicazione: (2023)
di: Chakraborty, Souradip, et al.
Pubblicazione: (2023)
DMCA: Dense Multi-agent Navigation using Attention and Communication
di: Arul, Senthil Hariharan, et al.
Pubblicazione: (2022)
di: Arul, Senthil Hariharan, et al.
Pubblicazione: (2022)
PIPER: Primitive-Informed Preference-based Hierarchical Reinforcement Learning via Hindsight Relabeling
di: Singh, Utsav, et al.
Pubblicazione: (2024)
di: Singh, Utsav, et al.
Pubblicazione: (2024)
TOPGN: Real-time Transparent Obstacle Detection using Lidar Point Cloud Intensity for Autonomous Robot Navigation
di: Weerakoon, Kasun, et al.
Pubblicazione: (2024)
di: Weerakoon, Kasun, et al.
Pubblicazione: (2024)
MOSU: Autonomous Long-range Robot Navigation with Multi-modal Scene Understanding
di: Liang, Jing, et al.
Pubblicazione: (2025)
di: Liang, Jing, et al.
Pubblicazione: (2025)
ViLAM: Distilling Vision-Language Reasoning into Attention Maps for Social Robot Navigation
di: Elnoor, Mohamed, et al.
Pubblicazione: (2025)
di: Elnoor, Mohamed, et al.
Pubblicazione: (2025)
HALO: Human Preference Aligned Offline Reward Learning for Robot Navigation
di: Seneviratne, Gershom, et al.
Pubblicazione: (2025)
di: Seneviratne, Gershom, et al.
Pubblicazione: (2025)
REBEL: Reward Regularization-Based Approach for Robotic Reinforcement Learning from Human Feedback
di: Chakraborty, Souradip, et al.
Pubblicazione: (2023)
di: Chakraborty, Souradip, et al.
Pubblicazione: (2023)
Beyond Joint Demonstrations: Personalized Expert Guidance for Efficient Multi-Agent Reinforcement Learning
di: Yu, Peihong, et al.
Pubblicazione: (2024)
di: Yu, Peihong, et al.
Pubblicazione: (2024)
LANCAR: Leveraging Language for Context-Aware Robot Locomotion in Unstructured Environments
di: Shek, Chak Lam, et al.
Pubblicazione: (2023)
di: Shek, Chak Lam, et al.
Pubblicazione: (2023)
DIPPER: Direct Preference Optimization to Accelerate Primitive-Enabled Hierarchical Reinforcement Learning
di: Singh, Utsav, et al.
Pubblicazione: (2024)
di: Singh, Utsav, et al.
Pubblicazione: (2024)
AMCO: Adaptive Multimodal Coupling of Vision and Proprioception for Quadruped Robot Navigation in Outdoor Environments
di: Elnoor, Mohamed, et al.
Pubblicazione: (2024)
di: Elnoor, Mohamed, et al.
Pubblicazione: (2024)
On the Vulnerability of LLM/VLM-Controlled Robotics
di: Wu, Xiyang, et al.
Pubblicazione: (2024)
di: Wu, Xiyang, et al.
Pubblicazione: (2024)
MaxMin-RLHF: Alignment with Diverse Human Preferences
di: Chakraborty, Souradip, et al.
Pubblicazione: (2024)
di: Chakraborty, Souradip, et al.
Pubblicazione: (2024)
Robot Navigation Using Physically Grounded Vision-Language Models in Outdoor Environments
di: Elnoor, Mohamed, et al.
Pubblicazione: (2024)
di: Elnoor, Mohamed, et al.
Pubblicazione: (2024)
BehAV: Behavioral Rule Guided Autonomy Using VLMs for Robot Navigation in Outdoor Scenes
di: Weerakoon, Kasun, et al.
Pubblicazione: (2024)
di: Weerakoon, Kasun, et al.
Pubblicazione: (2024)
CoNVOI: Context-aware Navigation using Vision Language Models in Outdoor and Indoor Environments
di: Sathyamoorthy, Adarsh Jagan, et al.
Pubblicazione: (2024)
di: Sathyamoorthy, Adarsh Jagan, et al.
Pubblicazione: (2024)
Deceptive Path Planning via Reinforcement Learning with Graph Neural Networks
di: Fatemi, Michael Y., et al.
Pubblicazione: (2024)
di: Fatemi, Michael Y., et al.
Pubblicazione: (2024)
Signal attenuation enables scalable decentralized multi-agent reinforcement learning over networks
di: Suttle, Wesley A, et al.
Pubblicazione: (2025)
di: Suttle, Wesley A, et al.
Pubblicazione: (2025)
Direct Preference Optimization for Primitive-Enabled Hierarchical RL: A Bilevel Approach
di: Singh, Utsav, et al.
Pubblicazione: (2024)
di: Singh, Utsav, et al.
Pubblicazione: (2024)
Safety Recovery in Reasoning Models Is Only a Few Early Steering Steps Away
di: Ghosal, Soumya Suvra, et al.
Pubblicazione: (2026)
di: Ghosal, Soumya Suvra, et al.
Pubblicazione: (2026)
SABER: A Stealthy Agentic Black-Box Attack Framework for Vision-Language-Action Models
di: Wu, Xiyang, et al.
Pubblicazione: (2026)
di: Wu, Xiyang, et al.
Pubblicazione: (2026)
SAIL: Self-Improving Efficient Online Alignment of Large Language Models
di: Ding, Mucong, et al.
Pubblicazione: (2024)
di: Ding, Mucong, et al.
Pubblicazione: (2024)
DR. Nav: Semantic-Geometric Representations for Proactive Dead-End Recovery and Navigation
di: Rajagopal, Vignesh, et al.
Pubblicazione: (2025)
di: Rajagopal, Vignesh, et al.
Pubblicazione: (2025)
Learning Multi-Robot Coordination through Locality-Based Factorized Multi-Agent Actor-Critic Algorithm
di: Shek, Chak Lam, et al.
Pubblicazione: (2025)
di: Shek, Chak Lam, et al.
Pubblicazione: (2025)
Beyond Text: Utilizing Vocal Cues to Improve Decision Making in LLMs for Robot Navigation Tasks
di: Sun, Xingpeng, et al.
Pubblicazione: (2024)
di: Sun, Xingpeng, et al.
Pubblicazione: (2024)
TrustNavGPT: Modeling Uncertainty to Improve Trustworthiness of Audio-Guided LLM-Based Robot Navigation
di: Sun, Xingpeng, et al.
Pubblicazione: (2024)
di: Sun, Xingpeng, et al.
Pubblicazione: (2024)
Achieving Zero Constraint Violation for Constrained Reinforcement Learning via Conservative Natural Policy Gradient Primal-Dual Algorithm
di: Bai, Qinbo, et al.
Pubblicazione: (2022)
di: Bai, Qinbo, et al.
Pubblicazione: (2022)
AI Cap-and-Trade: Efficiency Incentives for Accessibility and Sustainability
di: Bornstein, Marco, et al.
Pubblicazione: (2026)
di: Bornstein, Marco, et al.
Pubblicazione: (2026)
Bounded Rationality for LLMs: Satisficing Alignment at Inference-Time
di: Chehade, Mohamad, et al.
Pubblicazione: (2025)
di: Chehade, Mohamad, et al.
Pubblicazione: (2025)
Transfer Q Star: Principled Decoding for LLM Alignment
di: Chakraborty, Souradip, et al.
Pubblicazione: (2024)
di: Chakraborty, Souradip, et al.
Pubblicazione: (2024)
Value of Information-based Deceptive Path Planning Under Adversarial Interventions
di: Suttle, Wesley A., et al.
Pubblicazione: (2025)
di: Suttle, Wesley A., et al.
Pubblicazione: (2025)
CROSS-GAiT: Cross-Attention-Based Multimodal Representation Fusion for Parametric Gait Adaptation in Complex Terrains
di: Seneviratne, Gershom, et al.
Pubblicazione: (2024)
di: Seneviratne, Gershom, et al.
Pubblicazione: (2024)
TopoNav: Topological Navigation for Efficient Exploration in Sparse Reward Environments
di: Hossain, Jumman, et al.
Pubblicazione: (2024)
di: Hossain, Jumman, et al.
Pubblicazione: (2024)
ROVER: Regulator-Driven Robust Temporal Verification of Black-Box Robot Policies
di: Sakano, Kristy, et al.
Pubblicazione: (2025)
di: Sakano, Kristy, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Towards Global Optimality for Practical Average Reward Reinforcement Learning without Mixing Time Oracles
di: Patel, Bhrij, et al.
Pubblicazione: (2024) -
Multi-LLM QA with Embodied Exploration
di: Patel, Bhrij, et al.
Pubblicazione: (2024) -
Personalized Embodied Navigation for Portable Object Finding
di: Dorbala, Vishnu Sashank, et al.
Pubblicazione: (2024) -
Code Comprehension then Auditing for Unsupervised LLM Evaluation
di: Patel, Bhrij, et al.
Pubblicazione: (2024) -
ProNav: Proprioceptive Traversability Estimation for Legged Robot Navigation in Outdoor Environments
di: Elnoor, Mohamed, et al.
Pubblicazione: (2023)